Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Enabling Agents to Communicate Entirely in Latent Space
Zhuoyun Du, Runze Wang, Huiyu Bai +6
While natural language is the de facto communication medium for LLM-based agents, it presents a fundamental constraint. The process of downsampling rich, internal latent states int…
cs.LG2026
Linear Dynamics in the RLVR Training of Large Language Models
Tianle Wang, Jiayu Liu, Zhongyuan Wu +4
Reinforcement learning with verifiable rewards (RLVR) has driven significant performance gains in reasoning-oriented large language models (LLMs), yet its internal training dynamic…