1 paper
Ningzhi Tang, Chaoran Chen, Gelei Xu +5
AI coding agents increasingly act directly within software environments, yet existing analyses of their failures rely on benchmark trajectories that miss how developers actually ex…