3 papers
cs.CV2026
Training-Free Open-Vocabulary Visual Grounding for Remote Sensing Images and Videos
Ke Li, Di Wang, Yongshan Zhu +5
Remote sensing visual grounding (RSVG) aims to localize a referred target in a remote sensing image or video according to a natural language expression. Existing RSVG methods usual…
cs.CV2026
ProVG: Progressive Visual Grounding via Language Decoupling for Remote Sensing Imagery
Ke Li, Ting Wang, Di Wang +4
Remote sensing visual grounding (RSVG) aims to localize objects in remote sensing imagery according to natural language expressions. Previous methods typically rely on sentence-lev…
cs.CV2025
RTS-Mono: A Real-Time Self-Supervised Monocular Depth Estimation Method for Real-World Deployment
Zeyu Cheng, Tongfei Liu, Tao Lei +3
Depth information is crucial for autonomous driving and intelligent robot navigation. The simplicity and flexibility of self-supervised monocular depth estimation are conducive to…