MegaTTS3
python tts/infer_cli.py --input_wav 'assets/English_prompt.wav' --input_text '这条音频的发音标准一些了吗?' --output_dir ./gen --p_w 2.5 --t_w 2.5.
text → audio · made by ByteDance
Sold by
Nobody in the catalogue publishes a price for this yet.
About
MegaTTS3 — an audio model from ByteDance.
It takes text and returns audio. It was published in March 2025. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.
Every current figure
- Maker
- ByteDance
- Register
- model
- Takes
- text
- Returns
- audio
- Published
- March 2025
- Licence
- apache-2.0
Known as 1 name
ByteDance/MegaTTS3