Pass IndexThe State of AISign in

Qwen3-TTS-12Hz-1.7B-VoiceDesign

We release Qwen3-TTS, a series of powerful speech generation models developed by Qwen, offering comprehensive support for voice cloning, voice design, ultra-high-quality human-like speech generation, and natural language-based voice control.

text → audio · made by Alibaba

Sold by

Nobody in the catalogue publishes a price for this yet.

About

Qwen3-TTS-12Hz-1.7B-VoiceDesign — an audio model from Alibaba.

It takes text and returns audio. It was published in January 2026. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.

Every current figure

Maker
Alibaba
Register
model
Takes
text
Returns
audio
Published
January 2026
Parameters
1.9 billion · read from its own weights
Licence
apache-2.0

Known as 1 name

Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign