Step-Audio-TTS-3B
Step-Audio-TTS-3B represents the industry's first Text-to-Speech (TTS) model trained on a large-scale synthetic dataset utilizing the LLM-Chat paradigm.
text → audio · made by stepfun-ai
Sold by
Nobody in the catalogue publishes a price for this yet.
About
Step-Audio-TTS-3B — an audio model from stepfun-ai.
It takes text and returns audio. It was published in February 2025. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.
Every current figure
- Maker
- stepfun-ai
- Register
- model
- Takes
- text
- Returns
- audio
- Published
- February 2025
- Parameters
- 3.5 billion · read from its own weights
- Licence
- apache-2.0
Known as 1 name
stepfun-ai/Step-Audio-TTS-3B