1 paper
Li Lin, Wujun Xu, Weiwei Meng +3
Vision-language-action models (VLAs), which leverage the cognition of multimodal information to infer physical-world actions, provide a generalized solution for embodied AI applica…