Qwen3-TTS-12Hz-1.7B-VoiceDesign
We release Qwen3-TTS, a series of powerful speech generation models developed by Qwen, offering comprehensive support for voice cloning, voice design, ultra-high-quality human-like speech generation, and natural language-based voice control.
text → audio · made by Alibaba
Sold by
Nobody in the catalogue publishes a price for this yet.
About
Qwen3-TTS-12Hz-1.7B-VoiceDesign — an audio model from Alibaba.
It takes text and returns audio. It was published in January 2026. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.
Every current figure
- Maker
- Alibaba
- Register
- model
- Takes
- text
- Returns
- audio
- Published
- January 2026
- Parameters
- 1.9 billion · read from its own weights
- Licence
- apache-2.0
Known as 1 name
Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign