Lance tables of LibriSpeech utterances at 16 kHz with cached last-layer representations from frozen microsoft/wavlm-base.
This release is a precomputed teacher cache plus raw audio bytes, not a general-purpose speech benchmark split.
Local / mirrored Hub layout:
Directory
Content
train/
LibriSpeech 960 h train corpora: train-clean-100, train-clean-360, train-other-500.
eval/
LibriSpeech dev-clean.
Each of train/ and eval/ is a… See the full description on the dataset page:
https://huggingface.co/datasets/Alright7398/modded-distill-wavlm-base.