2 papers
cs.LG2026
Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking
Vaidehi Bagaria, Nikshep Grampurohit, Pulkit Verma
Reinforcement learning (RL) allows vision-language-action (VLA) policies to generalize beyond their training distribution by optimizing directly for task success, but post-training…
cs.AI2026
Recursive Belief Vision Language Action Models
Vaidehi Bagaria, Bijo Sebastian, Nirav Kumar Patel
Vision-language-action models must enable agents to execute long-horizon tasks under partial observability. However, most existing approaches remain observation-driven, relying on…