1 paper · 1 filter
Yifei Wei, Linqing Zhong, Yi Liu +4
Vision-Language-Action (VLA) models are a promising paradigm for generalist robotic manipulation by grounding high-level semantic instructions into executable physical actions. How…