2 papers
cs.SE2026
Improving LLM Code Generation via Requirement-Aware Curriculum Reinforcement Learning
Shouyu Yin, Zhao Tian, Junjie Chen +1
Code generation, which aims to automatically generate source code from given programming requirements, has the potential to substantially improve software development efficiency. W…
cs.LG2025
SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM
Xiaojiang Zhang, Jinghui Wang, Zifei Cheng +14
Recent advances of reasoning models, exemplified by OpenAI's o1 and DeepSeek's R1, highlight the significant potential of Reinforcement Learning (RL) to enhance the reasoning capab…