This model is a 100-speaker multispeaker model in the Matcha-TTS format/architecture.(trained with Japanese)
This model is replaced 10 qwen-character to chatterbox(common voice) character.
Chatterbox
I faild to confirm watermark because of technical probrom,but maybe chatterbox watermark is exist.
If you don't like the watermark, use qwen3-tts only version
My training data is created by Apache Licensed/mit model output.
https://huggingface.co/Qwen/Qwen3-TTS-12Hz-1.7B-Base