Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Information Capacity: Evaluating the Efficiency of Large Language Models via Text Compression
Cheng Yuan, Jiawei Shao, Xuelong Li
Recent years have witnessed the rapid advancements of large language models (LLMs) and their expanding applications, leading to soaring demands for computational resources. The wid…
cs.AI2025
ScRPO: From Errors to Insights
Lianrui Li, Dakuan Lu, Jiawei Shao +1
We introduce Self-correction Relative Policy Optimization (ScRPO), a novel reinforcement learning framework designed to empower large language models with advanced mathematical rea…