2 papers
cs.AI2026
COMPASS: Cognitive MCTS-Guided Process Alignment for Safe Search Agents
Wenkai Shen, Pengyang Zhou, Jiahe Xu +5
LLM-powered search agents enable multi-step reasoning and tool use. However, these capabilities introduce retrieval-induced safety degradation, as harmful intents may decompose int…
cs.CL2026
ConflictBench: Evaluating Human-AI Conflict via Interactive and Visually Grounded Environments
Weixiang Zhao, Haozhen Li, Yanyan Zhao +5
As large language models (LLMs) evolve into autonomous agents capable of acting in open-ended environments, ensuring behavioral alignment with human values becomes a critical safet…