1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Chaoqun Cui, Jing Huang, Shijing Wang +3
Reinforcement learning with verifiable rewards (RLVR) provides a promising pathway for continuously advancing GUI agents, yet existing reward modeling paradigms face complementary…