Gemma 4 31B
Open Gemma instruction model for efficient chat and self-hosted deployments
Context
262,144
Max output
262,144
Input / 1M
$0.14
Output / 1M
$0.40
Providers
5
01Profile
Profile
Released
2026-04-04
Last updated
2026-05-02
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text · image
Output
text
Reasoning
No
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
—
Cache write
—
03Pricing
Pricing across 5 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
12+ recent events since 2026-04-04
2026-09-14Price upollama-cloudinput / 1M—→ $0.14
2026-09-14Price upollama-cloudoutput / 1M—→ $0.40
2026-08-12Price cutnano-gptinput / 1M$0.45→ $0.40
2026-08-12Price cutnano-gptinput / 1M$0.45→ $0.40
2026-08-12ContexttinfoilContext window256,000→ new
2026-08-07Addedregolo-aioffering→ new
2026-07-30Addedqvacoffering→ new
2026-06-26Addedtinfoiloffering→ new
2026-06-03Addednano-gptoffering→ new
2026-06-03Addednano-gptoffering→ new
2026-04-23Contextollama-cloudMax output8,192→ 262,144
2026-04-08Addedollama-cloudoffering→ new