2 papers
cs.CV2025
CoT4Det: A Chain-of-Thought Framework for Perception-Oriented Vision-Language Tasks
Yu Qi, Yumeng Zhang, Chenting Gong +4
Large Vision-Language Models (LVLMs) have demonstrated remarkable success in a broad range of vision-language tasks, such as general visual question answering and optical character…
cs.CV2025
LDMapNet-U: An End-to-End System for City-Scale Lane-Level Map Updating
Deguo Xia, Weiming Zhang, Xiyan Liu +6
An up-to-date city-scale lane-level map is an indispensable infrastructure and a key enabling technology for ensuring the safety and user experience of autonomous driving systems.…