11 citations · 11 across the 3 of their papers we have counts for
1 paper · 1 filter
Shu Zhao, Tan Yu, Anbang Xu +3
Reasoning-augmented search agents such as Search-R1, trained via reinforcement learning with verifiable rewards (RLVR), demonstrate remarkable capabilities in multi-step informatio…