Pass IndexThe State of AISign in

Step-Audio-TTS-3B

Step-Audio-TTS-3B represents the industry's first Text-to-Speech (TTS) model trained on a large-scale synthetic dataset utilizing the LLM-Chat paradigm.

text → audio · made by stepfun-ai

Sold by

Nobody in the catalogue publishes a price for this yet.

About

Step-Audio-TTS-3B — an audio model from stepfun-ai.

It takes text and returns audio. It was published in February 2025. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.

Every current figure

Maker
stepfun-ai
Register
model
Takes
text
Returns
audio
Published
February 2025
Parameters
3.5 billion · read from its own weights
Licence
apache-2.0

Known as 1 name

stepfun-ai/Step-Audio-TTS-3B