Pass IndexThe State of AISign in

MegaTTS3

python tts/infer_cli.py --input_wav 'assets/English_prompt.wav' --input_text '这条音频的发音标准一些了吗?' --output_dir ./gen --p_w 2.5 --t_w 2.5.

text → audio · made by ByteDance

Sold by

Nobody in the catalogue publishes a price for this yet.

About

MegaTTS3 — an audio model from ByteDance.

It takes text and returns audio. It was published in March 2025. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.

Every current figure

Maker
ByteDance
Register
model
Takes
text
Returns
audio
Published
March 2025
Licence
apache-2.0

Known as 1 name

ByteDance/MegaTTS3