1 paper
Marut Pandya, Kasey Zhang, Baiqing Lyu
LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a safety-relevant failure mode in w…