TLDRocket
Sign in

Voxtral transcribes at the speed of sound.

Mistral AI

Mistral released Voxtral Transcribe 2, a pair of speech-to-text models including Voxtral Mini for batch processing and Voxtral Realtime for live applications, with the latter available as open-weights software. Voxtral Mini achieves 4% word error rate at $0.003 per minute, outperforming competitors like GPT-4o mini and Gemini 2.5 Flash while processing audio 3x faster than ElevenLabs' Scribe v2 at one-fifth the cost. The models support 13 languages with speaker diarization, context biasing, and configurable latency down to sub-200 milliseconds, enabling applications from live subtitling to voice agents and contact center automation.

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.