Misc,

The FLORES-101 Evaluation Benchmark for Low-Resource and Multilingual Machine Translation

N. Goyal, C. Gao, V. Chaudhary, P. Chen, G. Wenzek, D. Ju, S. Krishnan, M. Ranzato, F. Guzman, and A. Fan.
(2021)cite arxiv:2106.03193.

Abstract

One of the biggest challenges hindering progress in low-resource and multilingual machine translation is the lack of good evaluation benchmarks. Current evaluation benchmarks either lack good coverage of low-resource languages, consider only restricted domains, or are low quality because they are constructed using semi-automatic procedures. In this work, we introduce the FLORES-101 evaluation benchmark, consisting of 3001 sentences extracted from English Wikipedia and covering a variety of different topics and domains. These sentences have been translated in 101 languages by professional translators through a carefully controlled process. The resulting dataset enables better assessment of model quality on the long tail of low-resource languages, including the evaluation of many-to-many multilingual translation systems, as all translations are multilingually aligned. By publicly releasing such a high-quality and high-coverage dataset, we hope to foster progress in the machine translation community and beyond.

BibTeX key: goyal2021flores101
entry type: misc
year: 2021
url: http://arxiv.org/abs/2106.03193
note: cite arxiv:2106.03193

BibSonomy

The FLORES-101 Evaluation Benchmark for Low-Resource and Multilingual Machine Translation

Abstract

Tags

Users

Comments and Reviewsshow / hide

Cite this publication

More citation styles

search on