NVIDIA Nemotron Cascade 2
Nemotron model for efficient reasoning, coding, and specialized AI agents
Context
262,144
Max output
131,072
Input / 1M
$0.15
Output / 1M
$0.60
Providers
1
01Profile
Profile
Released
2025-12-01
Last updated
2025-12-01
Knowledge cutoff
2024-07
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 1 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
5+ recent events since 2025-12-01
2026-07-28Removedveniceofferingremoved→ new
2026-06-10ContextveniceContext window256,000→ new
2026-06-10ContextveniceMax output32,768→ new
2026-05-23Addedvultroffering→ new
2026-04-09Addedveniceoffering→ new