2 papers
cs.CL2025
PersonaMath: Boosting Mathematical Reasoning via Persona-Driven Data Augmentation
Jing Luo, Longze Chen, Run Luo +12
While closed-source Large Language Models (LLMs) demonstrate strong mathematical problem-solving abilities, open-source models still face challenges with such tasks. To bridge this…
cs.CL2025
Quantification of Large Language Model Distillation
Sunbowen Lee, Junting Zhou, Chang Ao +11
Model distillation is a fundamental technique in building large language models (LLMs), transferring knowledge from a teacher model to a student model. However, distillation can le…