1 paper · 1 filter
Mo Li, L. H. Xu, Qitai Tan +2
Large language model (LLM)-based coding agents achieve impressive results on controlled benchmarks yet routinely produce pull requests that real maintainers reject. The root cause…