14 citations · 23 across the 4 of their papers we have counts for
1 paper · 1 filter
Quanfu Yu, Xian Wu, Hao Xu +1
Vision-Language-Action (VLA) models augmented with world modeling represent a promising paradigm for end-to-end autonomous driving. While pixel-level future prediction enables fine…