2 papers
cs.CL2026
Why LLMs Give In: Conversational Factors and Reasoning Behind Medical Sycophancy
Kaike Ping, Buse Çarık, Caleb Wohn +3
Large language models can answer a medical question correctly and still abandon that answer when a user pushes back. We study this failure as medical sycophancy and ask when models…
cs.SE2026
The Devil Is in the Interface: Evaluating How Tool Architecture Shapes Coding Agent Behavior
Xiangzhe Xu, Hamidreza Saghir, Qianhui Wu +5
As large language models continue to improve, agentic systems are becoming increasingly important, and tools are a key design dimension because they determine how agents access inf…