This dataset contains the train-unseen split of AnimalSpeak Pseudovox. Each
example is a short, silence-trimmed, single-vocalization WAV clip plus compact
per-clip metadata. It does not include generated conversations, captions, QA
pairs, or MCQ answers.
Rows: 346,907
Shards: 18
Maximum rows per shard: 20,000
data-20k/train-*.tar: WebDataset-style shards containing
audio/<audio_name> WAV entries.
metadata.parquet: one row per… See the full description on the dataset page:
https://huggingface.co/datasets/EarthSpeciesProject/animalspeak-pseudovox.