Best Input Price
$0.00 / 1M
Speech generation model for controllable voice, narration, and audio delivery
Context Window
0 tokens
Reasoning
No
Tool Calling
No
Released
2026-04-16
Inference Providers (2)
| Provider | Model ID | Context | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|
| StepFun (Global) | stepaudio-2.5-tts | 0 | $0.00 | $0.00 | Docs ↗ |
| StepFun (China) | stepaudio-2.5-tts | 0 | $0.00 | $0.00 | Docs ↗ |