Mistral Small 3.1 24B Instruct
Efficient Mistral model for fast chat, extraction, and production assistants
Context
128,000
Max output
128,000
Input / 1M
$0.10
Output / 1M
$0.30
Providers
4
01Profile
Profile
Released
2025-03-18
Last updated
2025-04-15
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
No
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 4 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-03-18
2026-09-14Addednano-gptoffering→ new
2026-08-25ContextkiloMax output128,000→ 102,400
2026-08-25ContextopenrouterMax output128,000→ 102,400
2026-08-24Removedcloudflare-ai-gatewayofferingremoved→ new
2026-08-14Price upcloudflare-ai-gatewayinput / 1M$0.35→ $0.35
2026-08-14Price cutcloudflare-ai-gatewayoutput / 1M$0.56→ $0.56
2026-08-14Contextcloudflare-ai-gatewayMax output16,384→ 128,000
2026-08-02Price upkiloinput / 1M$0.35→ $0.35
2026-08-02Price cutkilooutput / 1M$0.56→ $0.56
2026-08-02ContextkiloMax output131,072→ 128,000
2026-05-20Addedcloudflare-workers-aioffering→ new
2026-05-18ContextopenrouterMax output8,192→ 128,000