Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign (Alibaba Cloud, Apache-2.0).Qwen3TTSForConditionalGeneration (qwen3_tts)Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesignmodel.safetensors (talker) plus a speech_tokenizer/ module (12 Hz codec), config and tokenizer files.qwen3_tts architecture and loads with transformers (>= 4.57). It is API-compatible with the upstream base — follow the inference recipe on the base model card Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign.Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign (Apache-2.0). See NOTICE for full attribution.