1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information
Yauwai Yim, Chunkit Chan, Tianyu Shi +4
Large language models (LLMs) have shown success in handling simple games with imperfect information and enabling multi-agent coordination, but their ability to facilitate practical…
cs.CL2024
AICoderEval: Improving AI Domain Code Generation of Large Language Models
Yinghui Xia, Yuyan Chen, Tianyu Shi +2
Automated code generation is a pivotal capability of large language models (LLMs). However, assessing this capability in real-world scenarios remains challenging. Previous methods…