Training-ready derivative of shangeth/expresso-mimi-codes, built specifically for style-conditioned TTS fine-tuning.
⚠️ License: CC-BY-NC-4.0 — non-commercial use only.
Merges read + conversational configs into a single flat dataset per split (matches the canonical schema other *-mimi-codes datasets use).
Drops 5 styles whose ASR transcripts are not reliably aligned to the… See the full description on the dataset page:
https://huggingface.co/datasets/shangeth/expresso-mimi-codes-tagged.