Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Student Guides Teacher: Weak-to-Strong Inference via Spectral Orthogonal Exploration
Dayu Wang, Jiaye Yang, Weikang Li +4
Large Language Models (LLMs) often suffer from ''Reasoning Collapse'' on challenging mathematical reasoning tasks, where stochastic sampling produces lexical variations of the same…
cs.AI2025
Reducing Cognitive Overhead in Tool Use via Multi-Small-Agent Reinforcement Learning
Dayu Wang, Jiaye Yang, Weikang Li +2
Recent advances in multi-agent systems highlight the potential of specialized small agents that collaborate via division of labor. Existing tool-integrated reasoning systems, howev…