embodied agents 1multimodal learning 1physical intelligence 1policy optimization 1reinforcement learning 1reward design 1robotics 1text-to-video alignment 1video generation 1world action models 1
From the 2 of 24 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Information Capacity: Evaluating the Efficiency of Large Language Models via Text Compression
Cheng Yuan, Jiawei Shao, Xuelong Li
Recent years have witnessed the rapid advancements of large language models (LLMs) and their expanding applications, leading to soaring demands for computational resources. The wid…
cs.AI2026
ScRPO: From Errors to Insights
Lianrui Li, Dakuan Lu, Jiawei Shao +1
We introduce Self-correction Relative Policy Optimization (ScRPO), a novel reinforcement learning framework designed to empower large language models with advanced mathematical rea…