activity
20242026
most citedCooper: Co-Optimizing Policy and Reward Models in Reinforcement Learning for Large Language Models

1 citations · 5 across the 31 of their papers we have counts for

collaborators
Showing 2026 · cs.HCShow all

Nothing from them under that filter.

Their other years and fields are still on the left.