GLM 5.2 Short Flex
Flagship GLM model for hybrid reasoning, coding, and agentic engineering
Context
199,984
Max output
32,000
Input / 1M
$0.94
Output / 1M
$2.92
Providers
1
01Profile
Profile
Released
2026-06-17
Last updated
2026-06-17
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 1 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
4+ recent events since 2026-06-17
2026-08-27Price upneuralwattinput / 1M$0.72→ $0.94
2026-08-27Price upneuralwattoutput / 1M$2.25→ $2.92
2026-08-27ContextneuralwattMax output199,984→ 32,000
2026-06-28Addedneuralwattoffering→ new