[!NOTE]
For full documentation, detailed benchmark results, and methodology, please refer to the GitHub Repository.
J-HARD-TTS-Eval is a benchmark designed to evaluate the robustness of autoregressive Japanese Text-To-Speech (TTS) models.
It focuses on specific failure modes such as stability in short sequences, repetition handling, and context completion.
You can easily load the dataset using the Hugging Face datasets… See the full description on the dataset page:
https://huggingface.co/datasets/Parakeet-Inc/J-HARD-TTS-Eval.