Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL
Simon Yu, Nicholas Tomlin, Marwa Abdulhai +7
Multi-agent reinforcement learning for human-AI interaction typically relies on a single large language model to simulate user behavior. We show that this approach systematically f…
cs.CL2026
Overconfident and Blind to Details: Fixing Prompt Insensitivity with Abductive Preference Learning
Yijin Ni, Simon Yu, Peng Qi
Vision and language models frequently ignore semantically critical input edits, defaulting to pretraining priors. For example, models will confidently assert a five-legged dog has…