1 paper · 1 filter
Siqi Wen, Shu Yang, Shaopeng Fu +3
Vision Language Action (VLA) models close the perception action loop by translating multimodal instructions into executable behaviors, but this very capability magnifies safety ris…