RIR-Mega-Speech is a large-scale reverberant speech corpus created by convolving LibriSpeech utterances with simulated room impulse responses (RIRs sampled from the RIR-Mega collection). Each reverberant utterance includes per-file acoustic metadata computed from the source RIR, enabling controlled analysis of reverberation effects on speech processing systems.
This dataset emphasizes transparency and reproducibility: acoustic metrics are… See the full description on the dataset page: https://huggingface.co/datasets/mandipgoswami/rir-mega-speech.