3 papers
cs.CV2025
A Coarse-to-Fine Approach to Multi-Modality 3D Occupancy Grounding
Zhan Shi, Song Wang, Junbo Chen +1
Visual grounding aims to identify objects or regions in a scene based on natural language descriptions, essential for spatially aware perception in autonomous driving. However, exi…
cs.CV2024
UdeerLID+: Integrating LiDAR, Image, and Relative Depth with Semi-Supervised
Tao Ni, Xin Zhan, Tao Luo +3
Road segmentation is a critical task for autonomous driving systems, requiring accurate and robust methods to classify road surfaces from various environmental data. Our work intro…
cs.CV2024
MV-DETR: Multi-modality indoor object detection by Multi-View DEtecton TRansformers
Zichao Dong, Yilin Zhang, Xufeng Huang +4
We introduce a novel MV-DETR pipeline which is effective while efficient transformer based detection method. Given input RGBD data, we notice that there are super strong pretrainin…