Showing 2025Show all
2 papers · 1 filter
cs.AI2025
Tool-Augmented Policy Optimization: Synergizing Reasoning and Adaptive Tool Use with Reinforcement Learning
Wenxun Wu, Yuanyang Li, Guhan Chen +2
Recent advances in large language models (LLMs) have popularized test-time scaling, where models generate additional reasoning tokens before producing final answers. These approach…
cs.LG2025
Divergence of Empirical Neural Tangent Kernel in Classification Problems
Zixiong Yu, Songtao Tian, Guhan Chen
This paper demonstrates that in classification problems, fully connected neural networks (FCNs) and residual neural networks (ResNets) cannot be approximated by kernel logistic reg…