GLM 4.6 Turbo
Fast variant of GLM 4.6 for general chat, coding, and analysis with improved latency and strong reasoning.
Context
204,800
Max output
131,072
Input / 1M
$1.00
Output / 1M
$3.00
Providers
1
01Profile
Profile
Released
2025-10-02
Last updated
2025-10-02
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
No
Tool calling
No
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2025-10-02
2026-09-03Contextnano-gptContext window200,000→ 204,800
2026-09-03Contextnano-gptMax output204,800→ 131,072
2026-09-03Contextnano-gptContext window200,000→ 204,800
2026-09-03Contextnano-gptMax output204,800→ 131,072
2026-06-03Addednano-gptoffering→ new
2026-06-03Addednano-gptoffering→ new
2026-02-18Removednano-gptofferingremoved→ new
2026-02-18Removednano-gptofferingremoved→ new
2026-02-17Addednano-gptoffering→ new
2026-02-17Addednano-gptoffering→ new
2025-11-08Removedchutesofferingremoved→ new
2025-10-09Addedchutesoffering→ new