Active

Holo3-35B-A3B

Compact GPT model for low-latency assistance and high-volume workloads

JSONEmbed chartCompare
Context
65,536
Max output
8,192
Input / 1M
$0.25
Output / 1M
$1.80
Providers
1
01Profile

Profile

Released
2024-01-01
Last updated
2024-01-01
Knowledge cutoff
Weights
Open
02Capabilities

Capabilities

Input
text · image
Output
text
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Attachments
Yes
Cache read
Cache write
03Pricing

Pricing across 2 providers

Per 1M tokens, USD. Sorted by input price. First-party row highlighted.

ProviderProvider model idInputOutputCache readContext
nano-gptholo3-35b-a3b$0.25$1.80$0.1365,536
nano-gptholo3-35b-a3b:thinkingthinking$0.25$1.80$0.1365,536
04History

Change history

4+ recent events since 2024-01-01

2026-09-03Contextnano-gptMax output65,536→ 8,192
2026-09-03Contextnano-gptMax output65,536→ 8,192
2026-06-03Addednano-gptoffering→ new
2026-06-03Addednano-gptoffering→ new
as of 2026-09-16
About ·Corrections ·Contact ·Privacy ·EN / 中文 / 繁體
HomeModelsLabsProvidersToolsRankingsChangesNews