Pass IndexThe State of AISign in

Qwen3-TTS-12Hz-0.6B-CustomVoice

Qwen3-TTS is a series of advanced multilingual, controllable, robust, and streaming text-to-speech models developed by the Qwen team.

text → audio · made by Alibaba

Sold by

Nobody in the catalogue publishes a price for this yet.

About

Qwen3-TTS-12Hz-0.6B-CustomVoice — an audio model from Alibaba.

It takes text and returns audio. It was published in January 2026. The catalogue files it under speak. Its weights are published under APACHE-2.0, so you may run it on your own machine, or buy it from whoever serves it cheapest.

Every current figure

Maker
Alibaba
Register
model
Takes
text
Returns
audio
Published
January 2026
Parameters
906 million · read from its own weights
Licence
apache-2.0

Known as 1 name

Qwen/Qwen3-TTS-12Hz-0.6B-CustomVoice