This repository contains the Raw Machine Translation Predictions generated for the paper: "Translation as a Scalable Proxy for Multilingual Evaluation" (Issaka et al., 2026).
If you are looking for the aggregated evaluation scores (LM-Eval + MT Metrics), please see our companion repository: 👉 Link to Benchmark Scores Repo.
The rapid proliferation of LLMs has created a… See the full description on the dataset page:
https://huggingface.co/datasets/marslabucla/translation-proxy-paper-translations.