papers
Publications (2)
cs.CV2026
LWDrive: Layer-Wise World-Model-Guided Vision-Language Model Planning for Autonomous Driving
Chen Yang, Yuhao Wei, Ze Xu +6
Vision-Language Models (VLMs) provide powerful semantic understanding and commonsense reasoning for End-to-End Autonomous Driving (E2E-AD) planning. However, trajectories directly…
cs.CV2025
Shift-Window Meets Dual Attention: A Multi-Model Architecture for Specular Highlight Removal
Tianci Huo, Lingfeng Qi, Yuhan Chen +6
Inevitable specular highlights in practical environments severely impair the visual performance, thus degrading the task effectiveness and efficiency. Although there exist consider…