Synthesized speech from 7 open-source TTS systems on Taiwan-Mandarin / Chinese-English
code-switch sentences, across 4 input conditions (raw, controlled, ensub, ensub_ctrl).
Single source of truth for the blind-test Space
and the project's GitHub Pages site.
//.wav — audio clips (16/24/48 kHz depending on model)
sentences.jsonl — the 25 quick-test sentences (text, bucket, entities)
clips.jsonl — per-clip metadata:… See the full description on the dataset page: https://huggingface.co/datasets/JacobLinCool/zh-tw-tts-comparison.