Processed ViSEC speaker audio generated for the Meddies ASR collection.
Contents
processed_audio_by_id/: 147 WAV files named by speaker id.
metadata.csv: per-speaker metadata with duration, clip count, emotion coverage, and source-duration summary.
Schema
metadata.csv contains:
speaker_id: integer speaker identifier.
output_path: relative path to the processed WAV file.
duration_seconds: duration of the processed audio file.… See the full description on the dataset page: https://huggingface.co/datasets/Meddies/ViSEC-processed.