2 papers
cs.SE2026
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation
Shuyin Ouyang, Zhaozhi Qian, Faroq AL-Tam +2
Reinforcement Learning (RL) is an important paradigm for aligning Diffusion Language Models (DLMs) toward functional correctness in code generation. However, these models often enc…
cs.CL2025
Increasing the Thinking Budget is Not All You Need
Ignacio Iacobacci, Zhaozhi Qian, Faroq AL-Tam +2
Recently, a new wave of thinking-capable Large Language Models has emerged, demonstrating exceptional capabilities across a wide range of reasoning benchmarks. Early studies have b…