Prepared SpokenWoZ training dataset adapted for Whisper-style DST (dialog state tracking) tasks. Each example contains an utterance, normalized text, and aligned audio (16 kHz).
Key metadata
Number of examples: 73,950
Total size on disk: ~9.08 GB
Splits: train, validation
validation examples: 7,284
Features:
text (string): original utterance text
normalized_text (string): normalized form of the… See the full description on the dataset page: https://huggingface.co/datasets/vendrkat/spokenwoz_dst.