Dataset Card for CA-DE Parallel Corpus
Dataset Summary
The CA-DE Parallel Corpus is a Catalan-German dataset of parallel sentences created to support Catalan in NLP tasks, specifically
Machine Translation.
Supported Tasks and Leaderboards
The dataset can be used to train Bilingual Machine Translation models between German and Catalan in any direction,
as well as Multilingual Machine Translation models.