Findings of the 2009 workshop on statistical machine translation

Chris Callison-Burch, Philipp Koehn, Christof Monz, Josh Schroeder

Research output: Chapter in Book/Report/Conference proceedingConference contribution


This paper presents the results of the WMT09 shared tasks, which included a translation task, a system combination task, and an evaluation task. We conducted a large-scale manual evaluation of 87 machine translation systems and 22 system combination entries. We used the ranking of these systems to measure how strongly automatic metrics correlate with human judgments of translation quality, for more than 20 metrics. We present a new evaluation technique whereby system output is edited and judged for correctness.
Original languageEnglish
Title of host publicationProceedings of the Fourth Workshop on Statistical Machine Translation
Place of PublicationStroudsburg, PA USA
PublisherAssociation for Computational Linguistics
Number of pages28
Publication statusPublished - 2009

Publication series

NameStatMT '09
PublisherAssociation for Computational Linguistics

Cite this