This is the biggest and most comprehensive Romanian - Aromanian parallel corpus.
The Aromanian counterpart is automatically converted to Cunia and DIARO standards
More details about its collection and how it was used for the AroTranslate system can be found in our paper.
If you find our work usefull, please cite:
@article{jerpelea2024dialectal,
title={Dialectal and Low-Resource Machine Translation for Aromanian},
author={Jerpelea, Alexandru-Iulius and Rădoi, Alina and Nisioi, Sergiu}… See the full description on the dataset page:
https://huggingface.co/datasets/aronlp/aromanian-romanian-MT-corpus.