Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs
Kevin Wang, Anna Thöni, Benjamin Kempinski +50
Large language models (LLMs) are increasingly deployed as interactive agents, yet their capacity for social and strategic reasoning over extended interaction remains poorly underst…
cs.AI2024
Demystifying MuZero Planning: Interpreting the Learned Model
Hung Guei, Yan-Ru Ju, Wei-Yu Chen +1
MuZero has achieved superhuman performance in various games by using a dynamics network to predict the environment dynamics for planning, without relying on simulators. However, th…