1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025
WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement Learning
Kuan Li, Zhongwang Zhang, Huifeng Yin +14
Transcending human cognitive limitations represents a critical frontier in LLM training. Proprietary agentic systems like DeepResearch have demonstrated superhuman capabilities on…
cs.CL2025★ 1 cited
EvolveSearch: An Iterative Self-Evolving Search Agent
Dingchu Zhang, Yida Zhao, Jialong Wu +8
The rapid advancement of large language models (LLMs) has transformed the landscape of agentic information seeking capabilities through the integration of tools such as search engi…