2 papers
cs.MA2026
Theory of Mind Guided Strategy Adaptation for Zero-Shot Coordination
Andrew Ni, Simon Stepputtis, Stefanos Nikolaidis +3
A central challenge in multi-agent reinforcement learning is enabling agents to adapt to previously unseen teammates in a zero-shot fashion. Prior work in zero-shot coordination of…
cs.CL2024
Assessing biomedical knowledge robustness in large language models by query-efficient sampling attacks
R. Patrick Xian, Alex J. Lee, Satvik Lolla +4
The increasing depth of parametric domain knowledge in large language models (LLMs) is fueling their rapid deployment in real-world applications. Understanding model vulnerabilitie…