1 paper
Xiaoyu Lin, Xinkai Yu, Ankit Aich +2
Large Language Models (LLMs), which simulate human users, are frequently employed to evaluate chatbots in applications such as tutoring and customer service. Effective evaluation n…