17 citations · 17 across the 1 of their papers we have counts for
1 paper · 1 filter
Daniel Licht, Cynthia Gao, Janice Lam +3
Obtaining meaningful quality scores for machine translation systems through human evaluation remains a challenge given the high variability between human evaluators, partly due to…