GLM 5.3 Flash Uncensored
GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
Context
1,048,576
Max output
32,768
Input / 1M
$0.20
Output / 1M
$0.80
Providers
1
01Profile
Profile
Released
2026-07-29
Last updated
2026-08-27
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text · image
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 1 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2026-07-29
2026-09-14Price cutnano-gptinput / 1M$0.35→ $0.20
2026-09-14Price cutnano-gptoutput / 1M$1.40→ $0.80
2026-09-08Price upnano-gptinput / 1M$0.07→ $0.25
2026-09-08Price upnano-gptinput / 1M$0.25→ $0.35
2026-09-08Price upnano-gptoutput / 1M$0.21→ $1.00
2026-09-08Price upnano-gptoutput / 1M$1.00→ $1.40
2026-09-07Price cutnano-gptinput / 1M$0.35→ $0.07
2026-09-07Price cutnano-gptoutput / 1M$1.40→ $0.21
2026-09-02Contextnano-gptContext window262,144→ 524,288
2026-09-02Contextnano-gptContext window524,288→ 1,048,576
2026-08-30Price upnano-gptinput / 1M$0.13→ $0.35
2026-08-30Price upnano-gptoutput / 1M$0.50→ $1.40