Voxtral Small 24B
Mistral AI · Providers EU
Sovereignty and hosting
This model is catalogued for reference. It is not hosted on LLM EU yet, and no first-party endpoint serves it here. Requests to it go to the provider directly, under that provider's own terms.
Summary
Voxtral Small is an Apache-2.0 audio-input model that transcribes and reasons over speech, with a 32k-token context that the vendor maps to roughly 30 minutes of audio for transcription and 40 minutes for understanding. It also works as a text model and supports automatic source-language detection. The card lists English, Spanish, French, Portuguese, Hindi, German, Dutch and Italian as its target languages.
Specifications
- Provider
- Mistral AI (FR)
- Context window
- 33k
- Parameters
- 24.26B
- Modalities
- audio, text
- License
- apache-2.0 · licence
- Open weights
- yes
- Release date
- 01 Jul 2025
- Tasks
- translation, summarization, multilingual-support, customer-support
Languages
Strong: en, es, fr, pt, de, nl, it
Adequate: hi
Languages
Strong
- en
- es
- fr
- pt
- de
- nl
- it
Adequate
- hi
Not verified: the provider has not documented this claim. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.
When to use it
Strengths
- open-weight speech transcription and audio question answering
- automatic source-language detection
- also usable as a text model
- Apache-2.0 licence
Limits
- 32k-token window limits audio to about 30 to 40 minutes per request
- language coverage is narrower than Whisper’s
- no evaluation has been run by LLM EU