activity
20182022
most citedTo Ship or Not to Ship: An Extensive Evaluation of Automatic Metrics for Machine Translation

82 citations · 87 across the 4 of their papers we have counts for

collaborators

5 papers

stat.AP20223 cited

Searching for a higher power in the human evaluation of MT

Johnny Tian-Zheng Wei, Tom Kocmi, Christian Federmann

In MT evaluation, pairwise comparisons are conducted to identify the better system. In conducting the comparison, the experimenter must allocate a budget to collect Direct Assessme…

cs.CL20211 cited

The JHU-Microsoft Submission for WMT21 Quality Estimation Shared Task

Shuoyang Ding, Marcin Junczys-Dowmunt, Matt Post +2

This paper presents the JHU-Microsoft joint submission for WMT 2021 quality estimation shared task. We only participate in Task 2 (post-editing effort estimation) of the shared tas…

cs.CL202182 cited

To Ship or Not to Ship: An Extensive Evaluation of Automatic Metrics for Machine Translation

Tom Kocmi, Christian Federmann, Roman Grundkiewicz +3

Automatic metrics are commonly used as the exclusive tool for declaring the superiority of one machine translation system's quality over another. The community choice of automatic…

cs.CL20211 cited

On User Interfaces for Large-Scale Document-Level Human Evaluation of Machine Translation Outputs

Roman Grundkiewicz, Marcin Junczys-Dowmunt, Christian Federmann +1

Recent studies emphasize the need of document context in human evaluation of machine translations, but little research has been done on the impact of user interfaces on annotator p…

cs.CL2018

Achieving Human Parity on Automatic Chinese to English News Translation

Hany Hassan, Anthony Aue, Chang Chen +21

Machine translation has made rapid advances in recent years. Millions of people are using it today in online translation systems and mobile applications in order to communicate acr…