Mistral Nemo
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
Context
131,072
Max output
128,000
Input / 1M
$0.02
Output / 1M
$0.03
Providers
6
01Profile
Profile
Released
2024-07-01
Last updated
2024-07-30
Knowledge cutoff
2024-07
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
No
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
Not listed
Cache write
Not listed
03Pricing
Pricing across 6 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2024-07-01
2026-09-15Price cutvercelinput / 1M$0.15→ $0.04
2026-09-15Price upverceloutput / 1M$0.15→ $0.17
2026-09-15ContextvercelContext window128,000→ 60,288
2026-09-15ContextvercelMax output128,000→ 16,000
2026-08-30Price cutkiloinput / 1M$0.17→ $0.02
2026-08-30Price cutkilooutput / 1M$0.17→ $0.03
2026-08-24Price upkiloinput / 1M$0.02→ $0.17
2026-08-24Price upkilooutput / 1M$0.03→ $0.17
2026-08-03Removedgithub-modelsofferingremoved→ new
2026-08-02Price cutkiloinput / 1M$0.02→ $0.02
2026-08-02Price cutkilooutput / 1M$0.04→ $0.03
2026-07-25Removedazureofferingremoved→ new
05Usage
Monthly token share
Rank #59 · 2026-09
0.1%