Voxtral Small (latest)
Instruct model with native audio input for speech understanding and tool use
Context
32,768
Max output
32,000
Input / 1M
$0.10
Output / 1M
$0.30
Providers
3
01Profile
Profile
Released
2025-07-15
Last updated
2025-07-15
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text · audio
Output
text
Reasoning
No
Tool calling
Yes
Structured output
No
Attachments
Yes
Cache read
Not listed
Cache write
Not listed
03Pricing
Pricing across 3 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
3+ recent events since 2025-07-15
2026-08-28Addededenaioffering→ new
2026-08-16Addedllmtroffering→ new
2026-08-02Addedmistraloffering→ new