2 papers
cs.CV2026
DIFTA-3D: Depth-Consistent Instance-Level Feature Transfer and Adaptation of DINOv3 for 3D Detection
Linman Wang, ZiFei Zhang, Chunran Zheng +2
RGB-D 3D instance detectors benefit from visual semantics, but the task-specific Faster R-CNN/ResNet branch used by IIFNet3D couples feature extraction to a separately trained 2D d…
cs.RO2026
SparseNav: Instruction-conditioned Sparse Semantic Perception for Training-Free Vision-Language Navigation
Quanhua Chen, Juhan Kang, Runfeng Lin +5
Map-based vision-language navigation (VLN) relies on persistent spatial representations to connect language understanding with geometric planning. However, acquiring semantics beyo…