Grok 4.1 Fast (Non-Reasoning)
Fast Grok model for responsive chat, reasoning, and tool-assisted work
Context
2,000,000
Max output
2,000,000
Input / 1M
$0.18
Output / 1M
$0.45
Providers
13
01Profile
Profile
Released
2025-11-17
Last updated
2026-01
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text · image
Output
text
Reasoning
No
Tool calling
Yes
Structured output
No
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 11 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-11-17
2026-09-10Addedgoogle-vertexoffering→ new
2026-09-08Removed302aiofferingremoved→ new
2026-09-06StatusazureStatusbeta→ new
2026-08-20Addedllmgateway-providersoffering→ new
2026-08-10DeprecatedzenmuxStatus—→ deprecated
2026-06-24Addedllmgatewayoffering→ new
2026-06-06ContextvercelContext window2,000,000→ 1,000,000
2026-06-06ContextvercelMax output30,000→ 1,000,000
2026-05-18ContextvercelContext window2,000,000→ 1,000,000
2026-05-18ContextvercelContext window1,000,000→ 2,000,000
2026-05-18ContextvercelMax output30,000→ 1,000,000
2026-05-18ContextvercelMax output1,000,000→ 30,000