Odia Diverse ENsemble Speech Corpus
ODEN‑speech merges eight publicly‑available Odia (ଓଡ଼ିଆ) speech corpora into a single 16 kHz, speaker‑aware, text‑cleaned dataset suitable for ASR, TTS, representation learning and multilingual research.
🗂️ Source
Hours
License
Mozilla Common Voice 17 (Odia)
110 h
MPL‑2.0
LibriTTS (clean + other)
170 h
CC‑BY‑4.0
LJSpeech 1.1
24 h
CC‑BY‑4.0
VCTK (Odia & misc.)
40 h
CC‑BY‑4.0
IndicTTS… See the full description on the dataset page:
https://huggingface.co/datasets/BBSRguy/ODEN-speech.