6 citations · 7 across the 12 of their papers we have counts for
1 paper · 1 filter
Bince Qu, Wanli Li, Bo Pan +5
Reinforcement Learning (RL) has emerged as a powerful training paradigm for LLM-based agents. However, scaling agentic RL for deep research remains constrained by two coupled chall…