1 paper
Shuai Huang, Wenxuan Zhao, Jun Gao
As large language models (LLMs) develop anthropomorphic abilities, they are increasingly being deployed as autonomous agents to interact with humans. However, evaluating their perf…