2 citations · 2 across the 3 of their papers we have counts for
3 papers
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation
Shuyin Ouyang, Zhaozhi Qian, Faroq AL-Tam +2
Reinforcement Learning (RL) is an important paradigm for aligning Diffusion Language Models (DLMs) toward functional correctness in code generation. However, these models often enc…
Increasing the Thinking Budget is Not All You Need
Ignacio Iacobacci, Zhaozhi Qian, Faroq AL-Tam +2
Recently, a new wave of thinking-capable Large Language Models has emerged, demonstrating exceptional capabilities across a wide range of reasoning benchmarks. Early studies have b…
CamelEval: Advancing Culturally Aligned Arabic Language Models and Benchmarks
Zhaozhi Qian, Faroq Altam, Muhammad Alqurishi +1
Large Language Models (LLMs) are the cornerstones of modern artificial intelligence systems. This paper introduces Juhaina, a Arabic-English bilingual LLM specifically designed to…