1 paper
Chenglong Wang, Hang Zhou, Kaiyan Chang +6
Automatic evaluation of sequence generation, traditionally reliant on metrics like BLEU and ROUGE, often fails to capture the semantic accuracy of generated text sequences due to t…