1 citations · 1 across the 13 of their papers we have counts for
1 paper · 1 filter
Chanuk Lee, Minki Kang, Sung Ju Hwang
Recent studies observe that reinforcement learning with verifiable rewards (RLVR) reliably improves pass@1 on reasoning tasks, yet often fails to yield comparable gains in pass@k,…