Qwen3-TTS

Qwen3-TTS

Voice design, cloning & 97ms streaming

Product screenshot

About

A family of SOTA speech models (0.6B & 1.7B) supporting 10 languages. Features prompt-based Voice Design, 3s zero-shot cloning, and extreme low-latency streaming.