Zhipu AIglm familyActive

GLM 4.7 Flash Heretic

Efficient GLM model for fast reasoning, coding, and agent workflows

JSONEmbed chartCompare
Context
200,000
Max output
24,000
Input / 1M
$0.07
Output / 1M
$0.40
Providers
1
01Profile

Profile

Released
2026-02-04
Last updated
2026-06-11
Knowledge cutoff
Weights
Open
02Capabilities

Capabilities

Input
text
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
No
Cache read
Cache write
03Pricing

Pricing across 1 providers

Per 1M tokens, USD. Sorted by input price. First-party row highlighted.

ProviderProvider model idInputOutputCache readContext
veniceolafangensan-glm-4.7-flash-heretic$0.07$0.40$0.04200,000
04History

Change history

9+ recent events since 2026-02-04

2026-07-03Price cutveniceinput / 1M$0.14→ $0.07
2026-07-03Price cutveniceoutput / 1M$0.80→ $0.40
2026-06-10ContextveniceContext window200,000→ new
2026-06-10ContextveniceContext window→ 200,000
2026-06-10ContextveniceMax output24,000→ new
2026-06-10ContextveniceMax output→ 24,000
2026-03-12ContextveniceContext window128,000→ 200,000
2026-03-12ContextveniceMax output32,000→ 24,000
2026-02-18Addedveniceoffering→ new
as of 2026-09-16
About ·Corrections ·Contact ·Privacy ·EN / 中文 / 繁體
HomeModelsLabsProvidersToolsRankingsChangesNews