1 paper · 1 filter
Zhixuan Liang, Yuxiao Chen, Yurong You +12
Monolithic vision-action models represent an emerging paradigm in autonomous driving. However, this architecture produces token sequences that quickly exceed real-time computationa…