conference-paper Open access

Has Machine Translation Achieved Human Parity? A Case for Document-level Evaluation

Research footprint

At a glance

Citations
283
References
27
Comments
0
Paper overview

Abstract

Recent research suggests that neural machine translation achieves parity with professional human translation on the WMT Chinese-English news translation task. We empirically test this claim with alternative evaluation protocols, contrasting the evaluation of single sentences and entire documents. In a pairwise ranking experiment, human raters assessing adequacy and fluency show a stronger preference for human over machine translation when evaluating documents as compared to isolated sentences. Our findings emphasise the need to shift towards document-level evaluation as machine translation improves to the degree that errors which are hard or impossible to spot at the sentence-level become decisive in discriminating quality of different translation outputs.

Record transparency

Publication details

DOI
10.18653/v1/d18-1512
OpenAlex
W2888159079
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.