Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
simchoir – Dataset by arda-argmax | AlphaNeural AI
You can deploy this model and start earning money today!
arda-argmax
/
simchoir
like
0
automatic-speech-recognition
voice-activity-detection
cc-by-4.0
us
diarization
speaker-diarization
multi-speaker
synthetic
lhotse
fastmss
Views
No views yet
Model card
Files and Versions
Community
API
FastMSS synthetic multi-speaker meetings
Synthetic multi-speaker conversational audio generated with FastMSS. Each split contains mixture WAVs (16 kHz, mono), Lhotse manifests (recordings / supervisions / cuts), and per-mixture RTTM files with word-level speaker labels.
Splits / subfolders
debug/ — 1 mixtures, 1.6 min total, 6 unique speakers v0.1/ — 1000 mixtures, 1546.0 min total, 40 unique speakers
Per-split layout
/ audio/<recording_id>.wav… See the full description on the dataset page:
https://huggingface.co/datasets/arda-argmax/simchoir
.