1 paper
Yuanhao Li, Hongbo Wang, Xiaotang Shang +3
Reinforcement learning for program repair is hindered by sparse execution feedback and coarse sequence-level rewards that obscure which edits actually fix bugs. We present BoostAPR…