3 papers
cs.CV2026
Last-Layer-Centric Feature Recombination: Unleashing 3D Geometric Knowledge in DINOv3 for Monocular Depth Estimation
Gongshu Wang, Zhirui Wang, Kan Yang
Monocular depth estimation (MDE) is a fundamental yet inherently ill-posed task. Recent vision foundation models (VFMs), particularly DINO-based transformers, have significantly im…
cs.CV2026
EIMC: Efficient Instance-aware Multi-modal Collaborative Perception
Kang Yang, Peng Wang, Lantao Li +4
Multi-modal collaborative perception calls for great attention to enhancing the safety of autonomous driving. However, current multi-modal approaches remain a ``local fusion to com…
cs.CV2025
Point-PRC: A Prompt Learning Based Regulation Framework for Generalizable Point Cloud Analysis
Hongyu Sun, Qiuhong Ke, Yongcai Wang +4
This paper investigates the 3D domain generalization (3DDG) ability of large 3D models based on prevalent prompt learning. Recent works demonstrate the performances of 3D point clo…