DRAL is a bilingual speech corpus of parallel utterances, using recorded conversations and fragments re-enacted in a different language. It is intended as a resource for research, especially for training and evaluating speech-to-speech translation models and systems. We dedicate this corpus to the public domain; there is no copyright (CC 0).
DRAL is described in a new technical report: Dialogs Re-enacted Across Languages, Version… See the full description on the dataset page:
https://huggingface.co/datasets/jonavila/DRAL.