1 citations · 1 across the 11 of their papers we have counts for
1 paper · 1 filter
Xinyu Wang, Mingze Li, Sicheng Lyu +6
Vision-Language-Action (VLA) models unify perception, reasoning, and control in a single policy, but their multi-billion-parameter backbones and diffusion-based action heads make o…