1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2025
Clapper: Compact Learning and Video Representation in VLMs
Lingyu Kong, Hongzhi Zhang, Jingyuan Zhang +4
Current vision-language models (VLMs) have demonstrated remarkable capabilities across diverse video understanding applications. Designing VLMs for video inputs requires effectivel…
cs.RO2025
DiffAD: A Unified Diffusion Modeling Approach for Autonomous Driving
Tao Wang, Cong Zhang, Xingguang Qu +3
End-to-end autonomous driving (E2E-AD) has rapidly emerged as a promising approach toward achieving full autonomy. However, existing E2E-AD systems typically adopt a traditional mu…
cs.RO2024★ 1 cited
Solving Motion Planning Tasks with a Scalable Generative Model
Yihan Hu, Siqi Chai, Zhening Yang +6
As autonomous driving systems being deployed to millions of vehicles, there is a pressing need of improving the system's scalability, safety and reducing the engineering cost. A re…