A sound events dataset with multimodal3 modality, stored in parquet format.
Preprocessing & Augmentation
Preprocessing: standard
Augmentation: mixup cutmix
Split strategy: temporal
Sampling: weighted
Quality filtering: adaptive
Labeling: self training
clean.py — main artifact of this repository
See the… See the full description on the dataset page:
https://huggingface.co/datasets/stefanschne/mistral-chat.