🏆 Leaderboard | 🛠️ Evaluation Suite
TTS Voice Direction is a benchmark of 700 reference-conditioned speech
generation tasks. It evaluates whether a text-to-speech model can preserve a
reference speaker while following a natural-language direction that controls
how a new transcript is performed.
The benchmark emphasizes practical voice acting beyond basic emotion control.
It covers accent, acoustic delivery, vocal events, emotion, physiological… See the full description on the dataset page:
https://huggingface.co/datasets/BreezeBlue/TTS-Voice-Direction-Benchmark.