OpenHQ-SpeechT-GL-EN is Galician-to-English dataset for Speech Translation task.
This dataset has been compiled from OpenSLR's Crowdsourced high-quality Galician speech data set. This repository contains this dataset, originaly created by Google Reserach Team. (Check citation for more info).
It contains ~10h20m of galician male and female audios along with its text transcriptions and the correspondant English translations.