A 100-prompt English subset of zhaochenyang20/seed-tts-eval
(which is itself a HuggingFace mirror of BytedanceSpeech/seed-tts-eval),
used by vllm-omni nightly perf CI for
Qwen3-TTS Base voice_clone evaluation.
en/
├── meta.lst # 100 lines, 4-field pipe-delimited:
│ # utt_id|ref_text|prompt_wav_rel|target_text
└── prompt-wavs/ # WAVs referenced by meta.lst[2]
The bench reads… See the full description on the dataset page:
https://huggingface.co/datasets/linyueqian/seed-tts-eval-subset.