22 citations · 22 across the 2 of their papers we have counts for
2 papers
cs.CL2026
DISPO: Enhancing Training Efficiency and Stability in Reinforcement Learning for Large Language Model Mathematical Reasoning
Batuhan K. Karaman, Aditya Rawal, Suhaila Shakiah +4
Reinforcement learning with verifiable rewards has emerged as a promising paradigm for enhancing the reasoning capabilities of large language models particularly in mathematics. Cu…
astro-ph.GA2020★ 22 cited
Aliphatic hydrocarbon content of the interstellar dust
B. Günay, T. W. Schmidt, M. G. Burton +5
In the interstellar medium, carbon is distributed between the gas and solid phases. However, while about half of the expected carbon abundance can be accounted for in the gas phase…