3 papers
cs.LG2026
Beyond Alignment: Expanding Reasoning Capacity via Manifold-Reshaping Policy Optimization
Dayu Wang, Jiaye Yang, Weikang Li +2
Reinforcement Learning with Verifiable Rewards (RLVR) has demonstrated remarkable success in enhancing the reasoning capabilities of Large Language Models (LLMs). However, recent s…
cs.AI2025
Reducing Cognitive Overhead in Tool Use via Multi-Small-Agent Reinforcement Learning
Dayu Wang, Jiaye Yang, Weikang Li +2
Recent advances in multi-agent systems highlight the potential of specialized small agents that collaborate via division of labor. Existing tool-integrated reasoning systems, howev…
cs.AR2025
CIMFlow: An Integrated Framework for Systematic Design and Evaluation of Digital CIM Architectures
Yingjie Qi, Jianlei Yang, Yiou Wang +6
Digital Compute-in-Memory (CIM) architectures have shown great promise in Deep Neural Network (DNN) acceleration by effectively addressing the "memory wall" bottleneck. However, th…