Voxtral Small

Best open-weight transcription model — also understands audio.

Voxtral Small builds on Mistral Small 3 with audio input, excelling at speech transcription, translation and audio understanding while keeping strong text performance. Apache 2.0 weights.

Provider
Mistral AI
Type
Speech-to-text
Released
Jul 15, 2025
Context window
33K tokens
Price
$0.10 input / $0.30 output per 1M tokens
Input
text, audio, file
Output
text
Open weights
Yes
License
Apache 2.0

Best for

Strengths

More from Mistral AI

Sources