1 paper
Quanfu Yu, Xian Wu, Hao Xu +1
Vision-Language-Action (VLA) models augmented with world modeling represent a promising paradigm for end-to-end autonomous driving. While pixel-level future prediction enables fine…