Open library

This portal has been archived. Explore the next generation of this technology.

Towards Efficient Machine Translation Evaluation by Modelling Annotators

lib:ddbcb4de9ace0e69 (v1.0.0)

Authors: Nitika Mathur,Timothy Baldwin,Trevor Cohn
Where published: ALTA 2018 12
Document: PDF DOI

Abstract URL: https://www.aclweb.org/anthology/U18-1010/

Accurate evaluation of translation has long been a difficult, yet important problem. Current evaluations use direct assessment (DA), based on crowd sourcing judgements from a large pool of workers, along with quality control checks, and a robust method for combining redundant judgements. In this paper we show that the quality control mechanism is overly conservative, which increases the time and expense of the evaluation. We propose a model that does not rely on a pre-processing step to filter workers and takes into account varying annotator reliabilities. Our model effectively weights each worker's scores based on the inferred precision of the worker, and is much more reliable than the mean of either the raw scores or the standardised scores. We also show that DA does not deliver on the promise of longitudinal evaluation, and propose redesigning the structure of the annotation tasks that can solve this problem.

Relevant initiatives

Related knowledge about this paper

Search on this portal

Reproduced results (crowd-benchmarking and competitions)

Artifact and reproducibility checklists

Common formats for research projects and shared artifacts

Collective Knowledge (organizing research projects based on FAIR principles)

Reproducibility initiatives

Comments

Please log in to add your comments!

If you notice any inapropriate content that should not be here, please report us as soon as possible and we will try to remove it within 48 hours!

Towards Efficient Machine Translation Evaluation by Modelling Annotators

Relevant initiatives Hide

Comments Hide

Relevant initiatives

Comments