2 papers
cs.CV2026
Training-Free Open-Vocabulary Visual Grounding for Remote Sensing Images and Videos
Ke Li, Di Wang, Yongshan Zhu +5
Remote sensing visual grounding (RSVG) aims to localize a referred target in a remote sensing image or video according to a natural language expression. Existing RSVG methods usual…
cs.CV2026
Are Dense Labels Always Necessary for 3D Object Detection from Point Cloud?
Chenqiang Gao, Chuandong Liu, Jun Shu +5
Current state-of-the-art (SOTA) 3D object detection methods often require a large amount of 3D bounding box annotations for training. However, collecting such large-scale densely-s…