Open library

This portal has been archived. Explore the next generation of this technology.

Using Mechanical Turk to Build Machine Translation Evaluation Sets

lib:2130666e745b3826 (v1.0.0)

Authors: Michael Bloodgood,Chris Callison-Burch
ArXiv: 1410.5491
Document: PDF DOI

Abstract URL: http://arxiv.org/abs/1410.5491v1

Building machine translation (MT) test sets is a relatively expensive task. As MT becomes increasingly desired for more and more language pairs and more and more domains, it becomes necessary to build test sets for each case. In this paper, we investigate using Amazon's Mechanical Turk (MTurk) to make MT test sets cheaply. We find that MTurk can be used to make test sets much cheaper than professionally-produced test sets. More importantly, in experiments with multiple MT systems, we find that the MTurk-produced test sets yield essentially the same conclusions regarding system performance as the professionally-produced test sets yield.

Relevant initiatives

Related knowledge about this paper

Search on this portal

Reproduced results (crowd-benchmarking and competitions)

Artifact and reproducibility checklists

Common formats for research projects and shared artifacts

Collective Knowledge (organizing research projects based on FAIR principles)

Reproducibility initiatives

Comments

Please log in to add your comments!

If you notice any inapropriate content that should not be here, please report us as soon as possible and we will try to remove it within 48 hours!

Using Mechanical Turk to Build Machine Translation Evaluation Sets

Relevant initiatives Hide

Comments Hide

Relevant initiatives

Comments