SNAC-tokenized, flattened + augmented build of 5 subsets of
laion/voiceclap-data
(CC-BY-4.0): emolia (English block), ears, expresso, voxceleb1,
voxceleb2. Replaces the earlier single-subset
voiceclap-emolia-flattened repo (consolidated here).
Follow-up to an ablation study (2/3/5) on
/ token format:
full 7-tok/frame content with a SEPARATE vocab (ids shifted
+1,000,000 vs 's identical SNAC codes)… See the full description on the dataset page: https://huggingface.co/datasets/EmpathicRobotics/voiceclap-flattened.