Views
No views yet
microsoft/speecht5_tts finetuned on
sapinsapin/pld.samples/ (speechbrain x-vector speaker conditioning + microsoft/speecht5_hifigan vocoder).| metric | value |
|---|---|
| eval_loss | 0.4070 |
finetune_tts.py from the
halohalo pipeline; the dataset
adapter normalizes each corpus to (audio@16k, text, speaker_id) so corpora
are swappable with a --dataset flag.