1 citations · 1 across the 1 of their papers we have counts for
1 paper
Yufan Ye, Ting Zhang, Wenbin Jiang +1
Existing reinforcement learning strategies based on outcome supervision have proven effective in enhancing the performance of large language models(LLMs) for code generation. While…