1 paper
Hui Yu, Xiaofeng Wu, Wenbin Jiang +2
The widely-used automatic evaluation metrics cannot adequately reflect the fluency of the translations. The n-gram-based metrics, like BLEU, limit the maximum length of matched fra…