attention mechanisms 1camouflaged object detection 1depth perception 1multi-modality alignment 1rgb-d fusion 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CV2026
VCP-DCN: Beyond Visual Concealed Property via Depth Collaborative Network for Camouflaged Object Detection
Songsong Duan, Xi Yang, Nannan Wang
The paper proposes VCP-DCN, a depth collaborative network that aligns, interacts, and fuses RGB and depth features to improve camouflaged object detection in complex scenes.
cs.CV2025
Unleashing the Multi-View Fusion Potential: Noise Correction in VLM for Open-Vocabulary 3D Scene Understanding
Xingyilang Yin, Jiale Wang, Xi Yang +3
Recent open-vocabulary 3D scene understanding approaches mainly focus on training 3D networks through contrastive learning with point-text pairs or by distilling 2D features into 3…