2 papers
cs.AI2026
CORE: Concept-Oriented Reinforcement for Bridging the Definition-Application Gap in Mathematical Reasoning
Zijun Gao, Zhikun Xu, Xiao Ye +1
Large language models (LLMs) often solve challenging math exercises yet fail to apply the concept right when the problem requires genuine understanding. Popular Reinforcement Learn…
stat.ML2025
Statistical Inference for Generative Model Comparison
Zijun Gao, Yan Sun, Han Su
Generative models have achieved remarkable success across a range of applications, yet their evaluation still lacks principled uncertainty quantification. In this paper, we develop…