This is a compilation of parallel corpora resources used in building of Machine Translation engines in NTEU project (Action number: 2018-EU-IA-0051). Data in this resource are compiled in two TMX files, two tiers grouped by data source reliablity. Tier A -- danta originating from human edited sources, translation memories and alike. Tier B -- danta originating created by automatic… See the full description on the dataset page:
https://huggingface.co/datasets/FrancophonIA/NTEU_French-Slovenian.