2 papers
cs.CL2026
EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments
Jundong Xu, Qingchuan Li, Jiaying Wu +11
Large language model (LLM) agents have achieved strong performance on a wide range of benchmarks, yet most evaluations assume static environments. In contrast, real-world deploymen…
cs.AI2025
DVM: Towards Controllable LLM Agents in Social Deduction Games
Zheng Zhang, Yihuai Lan, Yangsen Chen +3
Large Language Models (LLMs) have advanced the capability of game agents in social deduction games (SDGs). These games rely heavily on conversation-driven interactions and require…