A curated benchmark dataset of 22,008 Croatian parliamentary speech clips extracted from ParlaSpeech-HR v3. Each clip includes aligned audio (WAV) and TextGrid annotations for linguistic analysis.
22,008 audio segments (various durations)
17,622 clips with complete TextGrid triplets:
.align (word-level boundaries via WordAlign tier)
.stress (primary stress frame labels; derivative of .align)
.pause (filled pause annotations… See the full description on the dataset page:
https://huggingface.co/datasets/porupski/ParlaSpeech-HR-benchmark_v3.