Open audio-language model for speech transcription, audio understanding, and voice-driven tool use
1 provider
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open audio-language model for speech transcription, audio understanding, and voice-driven tool use
Instruct model with native audio input for speech understanding and tool use
Open audio-language model for speech transcription, audio understanding, and voice-driven tool use
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Voxtral Mini 3B 2507mistral/voxtral-mini-3b-2507 | 32.768K | $0.04 | $0.04 | 2025-07-15 | |||
| Voxtral Small (latest)mistral/voxtral-small-latest | 32K | $0.1 | $0.3 | 2025-07-15 | |||
| Voxtral Small 24B 2507mistral/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | 2025-07-15 | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — |