The same per-token trajectory data as
abotresol/emotion-combined-trajectories-gemma-4-31b-it-v2, packed as
WebDataset tars instead of
5,888 separate .npz files.
Nothing changed but the packaging. Each tar member is the original .npz,
byte for byte; the build script checks every sample's SHA-256 against the file
it came from before publishing. The original repository stays where it is.
Why: 5,888 files in one directory is… See the full description on the dataset page:
https://huggingface.co/datasets/abotresol/emotion-combined-trajectories-gemma-4-31b-it-v2-webdataset.