5 citations · 13 across the 20 of their papers we have counts for
1 paper · 1 filter
Pablo Valle, Sergio Segura, Shaukat Ali +1
Vision-Language-Action (VLA) models are multimodal robotic task controllers that, given an instruction and visual inputs, produce a sequence of low-level control actions (or motor…