2 papers
cs.LG2026
When and Why Does Unsupervised RL Succeed in Mathematical Reasoning? A Manifold Envelopment Perspective
Zelin Zhang, Fei Cheng, Chenhui Chu
Although outcome-based reinforcement learning (RL) significantly advances the mathematical reasoning capabilities of Large Language Models (LLMs), its reliance on computationally e…
cs.CL2025
Assessing Agentic Large Language Models in Multilingual National Bias
Qianying Liu, Katrina Qiyao Wang, Fei Cheng +1
Large Language Models have garnered significant attention for their capabilities in multilingual natural language processing, while studies on risks associated with cross biases ar…