2 papers
cs.AI2026
MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution
Aliaksei Korshuk, Alexander Buyantuev, Ilya Makarov
Training language model agents for multi-agent strategic interaction presents a core difficulty: the quality of any action may depend on future events that never materialize, on mo…
cs.AI2025
Evaluating Generalization Capabilities of LLM-Based Agents in Mixed-Motive Scenarios Using Concordia
Chandler Smith, Marwa Abdulhai, Manfred Diaz +83
Large Language Model (LLM) agents have demonstrated impressive capabilities for social interaction and are increasingly being deployed in situations where they might engage with bo…