1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Yufan Ye, Ting Zhang, Wenbin Jiang +1
Existing reinforcement learning strategies based on outcome supervision have proven effective in enhancing the performance of large language models(LLMs) for code generation. While…