Training dataset for the ParaSpeechCLAP-Situational and ParaSpeechCLAP-Combined models, from the paper:
ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining
Anuj Diwan, Eunsol Choi, David Harwath
Under review
This dataset contains the situational-tag subset of ParaSpeechCaps, filtered to include only examples annotated with situational (utterance-level) style tags. It is used to train the… See the full description on the dataset page:
https://huggingface.co/datasets/ajd12342/paraspeechcaps-situational-train.