Need a voice model for your domain? Trelis builds custom ASR, TTS, and voice agent pipelines for specialist verticals (legal, medical, finance, construction) and low-resource languages. Enquire or book a consultation →
A 50-clip benchmark for 2-speaker overlapping speech recognition, derived from the AMI Meeting Corpus test split.
Each clip is 8–28 seconds of real conversational meeting audio reconstructed as a 2-speaker virtual meeting, with separate… See the full description on the dataset page:
https://huggingface.co/datasets/Trelis/ami-2speaker-test.