GLM-4 32B (0414-128k)
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
Context
128,000
Max output
128,000
Input / 1M
$0.10
Output / 1M
$0.10
Providers
2
01Profile
Profile
Released
2025-04-14
Last updated
2025-04-14
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
No
Tool calling
Yes
Structured output
Yes
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
3+ recent events since 2025-04-14
2026-08-20Addedllmgateway-providersoffering→ new
2026-01-23ContextllmgatewayMax output—→ 16,384
2026-01-23Addedllmgatewayoffering→ new