1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2026
CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning
Shijie Zhang, Zheng Xiao, Shiyu Liu +7
Online reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning abilities of large language models, but most methods still…
hep-ex2026★ 1 cited
Evidence of transverse polarization of hyperon in
BESIII Collaboration, M. Ablikim, M. N. Achasov +684
Using events collected with the BESIII detector at the BEPCII collider, we report an evidence of transverse polarization with a si…