1 citations · 1 across the 8 of their papers we have counts for
1 paper · 1 filter
Xiaoyu Ma, Zhengqing Yuan, Zheyuan Zhang +3
Vision-language-action (VLA) models enable impressive zero shot manipulation, but their inference stacks are often too heavy for responsive web demos or high frequency robot contro…