Voxtral
Model ● Covered in 4 stories + Follow
Voxtral is an open-source speech understanding model released by Mistral, available in 24B and 3B variants that transcribe audio and answer questions about it across 9 languages with 70ms latency. The model is priced competitively at less than half the cost of comparable services like OpenAI Whisper, starting at $0.001 per minute, and is integrated into Mistral's Le Chat platform for voice interaction capabilities. Voxtral is being used in applications like Agent Draw, which enables voice-to-drawing generation on canvas interfaces.
Updated 8 August 2026
Specifications
No specifications recorded yet.
Latest developments
July 2026
April 2026
Multiple AI companies release new model variants and agent systems in March 2026 Model release