This model is a fine-tuned version of
facebook/mbart-large-50 on the dataset "Thermostatic/texts_parallel_corpus_europarl_english_spanish".
It achieves the following results on the evaluation set:
This small model was developed to translate the minutes of the European Parliament proceedings from English into Spanish as part of a Master’s degree project in Natural Language Processing (NLP).
The model was trained on a relatively small dataset (3,000 rows), so its performance is limited. It is intended primarily for educational and experimental purposes.
The model was trained on this dataset, which includes a large parallel corpus of English–Spanish sentence pairs:
https://huggingface.co/datasets/Thermostatic/texts_parallel_corpus_europarl_english_spanish