GLM 4.7 Flash Original
GLM-4.7-Flash is a lightweight 30B model optimized for coding and agentic tasks. Balances high performance with efficiency, perfect for local deployment.
Context
200,000
Max output
128,000
Input / 1M
$0.07
Output / 1M
$0.40
Providers
1
01Profile
Profile
Released
2026-01-19
Last updated
2026-01-19
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 2 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
6+ recent events since 2026-01-19
2026-06-03Addednano-gptoffering→ new
2026-06-03Addednano-gptoffering→ new
2026-02-18Removednano-gptofferingremoved→ new
2026-02-18Removednano-gptofferingremoved→ new
2026-02-17Addednano-gptoffering→ new
2026-02-17Addednano-gptoffering→ new