4 papers
SidConArena: An Environment Evaluating Agents in Open-Ended,Positive-Sum Bargaining Game
Yeqi Feng, Yuxin Chen, Tianxing He
Evaluating LLM agents requires dynamic environments that go beyond static reasoning and zero-sum games. Real-world economic interaction is often open-ended and mixed-motive: agents…
SimCity: Multi-Agent Urban Development Simulation with Rich Interactions
Yeqi Feng, Yucheng Lu, Hongyu Su +2
Large Language Models (LLMs) open new possibilities for constructing realistic and interpretable macroeconomic simulations. We present SimCity, a multi-agent framework that leverag…
Can LLMs Generate Reliable Test Case Generators? A Study on Competition-Level Programming Problems
Yuhan Cao, Zian Chen, Kun Quan +17
Large Language Models (LLMs) have demonstrated remarkable capabilities in code generation, capable of tackling complex tasks during inference. However, the extent to which LLMs can…
LLMs vs. Chinese Anime Enthusiasts: A Comparative Study on Emotionally Supportive Role-Playing
Lanlan Qiu, Xiao Pu, Yeqi Feng +1
Large Language Models (LLMs) have demonstrated impressive capabilities in role-playing conversations and providing emotional support as separate research directions. However, there…