Cleaned TTS dataset for Rongmei (nbu), a North East Indian language. Derived from the Vaani dataset with SNR filtering, LUFS normalization, and text cleaning.
Metric
Value
Total clips
193
Total hours
0.3h
High SNR (>=20dB)
127 clips
Medium SNR (15-20dB)
66 clips
Sample rate
16kHz
Audio format
WAV, 16-bit PCM
Column
Type
Description
audio
Audio
16kHz WAV audio
text
string… See the full description on the dataset page:
https://huggingface.co/datasets/sulabhkatiyar/ne-tts-nbu.