Dataset Card for EU Acts: Ukrainian → EU Languages (le-llm/EU-acts-uk-to-langs)
Dataset Description
Dataset Summary
EU-acts-uk-to-langs is a sentence- and segment-aligned parallel corpus of EU legal acts, pairing Ukrainian (uk) source segments with their counterparts in the 24 official EU languages. The corpus is derived from TMX memories built from: (a) translations of EU acts published at the official portal of the Parliament of Ukraine, and (b) EU acts available on… See the full description on the dataset page: https://huggingface.co/datasets/lapa-llm/EU-acts-uk-to-langs.