Llama 3.3 Nemotron Super 49B v1.5
Nemotron model for efficient reasoning, coding, and specialized AI agents
Context
131,072
Max output
131,072
Input / 1M
$0.40
Output / 1M
$0.40
Providers
2
01Profile
Profile
Released
2025-07-25
Last updated
2025-07-25
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 1 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-07-25
2026-08-03Removednano-gptofferingremoved→ new
2026-08-02Removedkiloofferingremoved→ new
2026-07-27Addednvidiaoffering→ new
2026-07-18DeprecateddeepinfraStatus—→ deprecated
2026-07-17Removedopenrouterofferingremoved→ new
2026-06-26Addeddeepinfraoffering→ new
2026-06-09Removednvidiaofferingremoved→ new
2026-06-07Price upopenrouterinput / 1M$0.10→ $0.40
2026-06-07ContextopenrouterContext window131,072→ new
2026-05-15Addedopenrouteroffering→ new
2026-02-16Addedkilooffering→ new
2025-12-27Addednano-gptoffering→ new