Contrastive test set for English-to-French MT evaluation covering 2 discourse phenomena: anaphora and lexical choice (coherence/cohesion).
Dataset Description
For machine translation to tackle discourse phenomena, models must have access to extra-sentential linguistic context. There has been recent interest in modelling context in neural machine translation, but models have been principally evaluated with standard… See the full description on the dataset page: https://huggingface.co/datasets/rbawden/DiscEvalMT.