The EA-MT dataset
- Read more about The EA-MT dataset
- Log in or register to post comments
The dataset used in this task is composed of source-language (English) sentences that contain named entities that are potentially complex from a machine translation perspective. These entities may be rare, ambiguous, or unknown to translation systems, posing an additional challenge beyond conventional lexical translation. The goal of the dataset is to evaluate the ability of machine translation systems to correctly handle such elements, ensuring their accurate transfer into the target language without loss of meaning or disambiguation errors.

