9 citations · 12 across the 8 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.AI2024
Exploring the Interplay Between Video Generation and World Models in Autonomous Driving: A Survey
Ao Fu, Yi Zhou, Tao Zhou +5
World models and video generation are pivotal technologies in the domain of autonomous driving, each playing a critical role in enhancing the robustness and reliability of autonomo…
cs.CV2024★ 1 cited
MePT: Multi-Representation Guided Prompt Tuning for Vision-Language Model
Xinyang Wang, Yi Yang, Minfeng Zhu +3
Recent advancements in pre-trained Vision-Language Models (VLMs) have highlighted the significant potential of prompt tuning for adapting these models to a wide range of downstream…
cs.CV2024
Delving into Multi-modal Multi-task Foundation Models for Road Scene Understanding: From Learning Paradigm Perspectives
Sheng Luo, Wei Chen, Wanxin Tian +12
Foundation models have indeed made a profound impact on various fields, emerging as pivotal components that significantly shape the capabilities of intelligent systems. In the cont…