Step TTS 2
Speech generation model for controllable voice, narration, and audio delivery
Context
0
Max output
0
Input / 1M
Not listed
Output / 1M
Not listed
Providers
2
01Profile
Profile
Released
2026-03-01
Last updated
2026-07-02
Knowledge cutoff
—
Weights
Closed
02Capabilities
Capabilities
Input
text
Output
audio
Reasoning
No
Tool calling
No
Structured output
No
Attachments
No
Cache read
—
Cache write
—
03Pricing
Pricing across 0 providers
Per 1M tokens, USD. Sorted by input price. First-party row highlighted.
04History
Change history
1+ recent events since 2026-03-01
2026-07-02Addedstepfunoffering→ new