Pith. sign in

REVIEW 1 cited by

The Effect of Translationese in Machine Translation Test Sets

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1906.08069 v1 pith:P5JD6KN6 submitted 2019-06-19 cs.CL

classification cs.CL
keywords translationtranslationesetesteffectsetsdatadirectionmachine
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

The effect of translationese has been studied in the field of machine translation (MT), mostly with respect to training data. We study in depth the effect of translationese on test data, using the test sets from the last three editions of WMT's news shared task, containing 17 translation directions. We show evidence that (i) the use of translationese in test sets results in inflated human evaluation scores for MT systems; (ii) in some cases system rankings do change and (iii) the impact translationese has on a translation direction is inversely correlated to the translation quality attainable by state-of-the-art MT systems for that direction.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. On The Evaluation of Machine Translation Systems Trained With Back-Translation

    cs.CL 2019-08 conditional novelty 6.0 of 10

    Back-translation produces more fluent, human-preferred output even when BLEU is flat, so evaluation should combine BLEU with a language model score.

Pith tools