Mistral AI has released Voxtral Transcribe 2, a next-generation speech-to-text model with state-of-the-art transcription quality, diarization, and ultra-low latency. The model comes in two versions: Voxtral Mini Transcribe V2 for batch transcription and Voxtral Realtime for live applications. Voxtral Realtime is open-source and available under the Apache 2.0 license. The models support 13 languages and offer industry-leading accuracy at a fraction of the cost of competitors.