2 papers
cs.CL2026
ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents
Xing Fu, Yulin Hu, Mengtong Ji +5
Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to users' latent emotional needs…
cs.CL2026
ConflictBench: Evaluating Human-AI Conflict via Interactive and Visually Grounded Environments
Weixiang Zhao, Haozhen Li, Yanyan Zhao +5
As large language models (LLMs) evolve into autonomous agents capable of acting in open-ended environments, ensuring behavioral alignment with human values becomes a critical safet…