GLM 4.6V Original
GLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).
Context
128,000
Max output
24,000
Input / 1M
$0.60
Output / 1M
$0.90
Providers
1
01Profile
Profile
Released
2025-12-08
Last updated
2025-12-08
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text · image
Output
text
Reasoning
No
Tool calling
No
Structured output
No
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 1 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
3+ recent events since 2025-12-08
2026-06-03Addednano-gptoffering→ new
2026-02-18Removednano-gptofferingremoved→ new
2026-02-17Addednano-gptoffering→ new