task_categories:
- audio-classification
language:
- en
tags:
- speaker-diarization
- test-dataset
size_categories:
- n<1K
Human-recorded meeting audio with ground truth speaker annotations for acceptance testing.
Version: 1.0.0
Clips: 3 meetings (2-4 minutes each)
Speakers: 2-4 per clip
Format: 16kHz mono WAV
Annotation: RTTM format (Rich Transcription Time Marked)… See the full description on the dataset page:
https://huggingface.co/datasets/zant-os/zant-echo-golden.