1 paper
Hengyi Xie, Chenfei Yao, Xianjin Wu +4
Vision-language-action (VLA) models commonly adopt an LLM-centric V→L→A pathway, processing visual observations and language instructions through a large language model b…