319,765 crossfaded speech trajectories · 4,482 audio-hours · 1,598,825 source clips
A trajectory is a short sequence of 5 consecutive utterances by one
speaker whose measured emotion or voice character moves monotonically from one end of
the corpus distribution to the other. The clips are joined into one continuous audio file
with equal-power crossfades, the joined audio is re-tokenized with MOSS-Audio-
Tokenizer-v2, and every… See the full description on the dataset page:
https://huggingface.co/datasets/laion/laion-emotional-trajectory-t80.