4 papers · 1 filter
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training
Chen Wang, Hexuan Deng, Yining Zhang +5
Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning traces. Existing methods ma…
Targeted Exploration via Unified Entropy Control for Reinforcement Learning
Chen Wang, Lai Wei, Yanzhi Zhang +5
Recent advances in reinforcement learning (RL) have improved the reasoning capabilities of large language models (LLMs) and vision-language models (VLMs). However, the widely used…
How Modality Shapes Perception and Reasoning: A Study of Error Propagation in ARC-AGI
Bo Wen, Chen Wang, Erhan Bilal
ARC-AGI and ARC-AGI-2 measure generalization-through-composition on small color-quantized grids, and their prize competitions make progress on these harder held-out tasks a meaning…
Voice-based AI Agents: Filling the Economic Gaps in Digital Health Delivery
Bo Wen, Chen Wang, Qiwei Han +4
The integration of voice-based AI agents in healthcare presents a transformative opportunity to bridge economic and accessibility gaps in digital health delivery. This paper explor…