NVIDIAnemotron familyActive

Llama 3.1 Nemotron 70B Instruct

Nemotron model for efficient reasoning, coding, and specialized AI agents

JSONEmbed chartCompare
Context
131,072
Max output
8,192
Input / 1M
$0.60
Output / 1M
$0.60
Providers
2
01Profile

Profile

Released
2025-04-15
Last updated
2025-04-15
Knowledge cutoff
Weights
Open
02Capabilities

Capabilities

Input
text
Output
text
Reasoning
No
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
Cache write
03Pricing

Pricing across 1 providers

Per 1M tokens, USD. Sorted by input price. First-party row highlighted.

ProviderProvider model idInputOutputCache readContext
nvidiafirst-partynvidia/llama-3.1-nemotron-70b-instructNot listedNot listedNot listed128,000
edenaideepinfra/nvidia/Llama-3.1-Nemotron-70B-Instruct$0.60$0.60Not listed131,072
04History

Change history

6+ recent events since 2025-04-15

2026-08-14Addededenaioffering→ new
2026-07-27Addednvidiaoffering→ new
2026-05-16Removedkiloofferingremoved→ new
2026-05-03Removednvidiaofferingremoved→ new
2026-02-16Addedkilooffering→ new
2025-12-24Addednvidiaoffering→ new
as of 2026-09-16
About ·Corrections ·Contact ·Privacy ·EN / 中文 / 繁體
HomeModelsLabsProvidersToolsRankingsChangesNews