2 papers
cs.AI2025
Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning
Jiayu Wang, Yifei Ming, Zixuan Ke +4
Reinforcement learning (RL) has become the dominant paradigm for improving the performance of language models on complex reasoning tasks. Despite the substantial empirical gains de…
cs.LG2025
COSMOS: Predictable and Cost-Effective Adaptation of LLMs
Jiayu Wang, Aws Albarghouthi, Frederic Sala
Large language models (LLMs) achieve remarkable performance across numerous tasks by using a diverse array of adaptation strategies. However, optimally selecting a model and adapta…