MiMo-V2-TTS
Speech generation model for controllable voice, narration, and audio delivery
Context
8,192
Max output
8,192
Input / 1M
Not listed
Output / 1M
Not listed
Providers
3
01Profile
Profile
Released
2026-03-18
Last updated
2026-03-18
Knowledge cutoff
—
Weights
Open
02Capabilities
Capabilities
Input
text
Output
audio
Reasoning
No
Tool calling
No
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 0 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
4+ recent events since 2026-03-18
2026-05-27Contextxiaomi-token-plan-cnMax output16,384→ 8,192
2026-05-06Contextxiaomi-token-plan-cnContext window8,000→ 8,192
2026-05-06Contextxiaomi-token-plan-cnMax output16,000→ 16,384
2026-04-03Addedxiaomi-token-plan-cnoffering→ new