1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2026
Beyond Stochastic Exploration: What Makes Training Data Valuable for Agentic Search
Chuzhan Hao, Wenfeng Feng, Guochao Jiang +3
Reinforcement learning (RL) has become an effective approach for advancing the reasoning capabilities of large language models (LLMs) through the strategic integration of external…
cs.CL2024★ 1 cited
Mixture-of-LoRAs: An Efficient Multitask Tuning for Large Language Models
Wenfeng Feng, Chuzhan Hao, Yuewei Zhang +2
Instruction Tuning has the potential to stimulate or enhance specific capabilities of large language models (LLMs). However, achieving the right balance of data is crucial to preve…