1 paper
Jingtao He, Hongliang Lu, Xiaoyun Qiu +2
Vision-Language-Action (VLA) models have demonstrated promising capability in autonomous driving, highlighting the potential of unified multimodal architectures for jointly modelin…