2 citations · 2 across the 1 of their papers we have counts for
3 papers
cs.CL2025★ 2 cited
A Survey of Reinforcement Learning for Large Reasoning Models
Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36
In this paper, we survey recent advances in Reinforcement Learning (RL) for reasoning with Large Language Models (LLMs). RL has achieved remarkable success in advancing the frontie…
math.OC2024
Local Linear Convergence of Infeasible Optimization with Orthogonal Constraints
Youbang Sun, Shixiang Chen, Alfredo Garcia +1
Many classical and modern machine learning algorithms require solving optimization tasks under orthogonality constraints. Solving these tasks with feasible methods requires a gradi…
cs.LG2023★ 2 cited
Provably Fast Convergence of Independent Natural Policy Gradient for Markov Potential Games
Youbang Sun, Tao Liu, Ruida Zhou +2
This work studies an independent natural policy gradient (NPG) algorithm for the multi-agent reinforcement learning problem in Markov potential games. It is shown that, under mild…