1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Lightweight Adaptation of General-Purpose VLMs for Multispectral and SAR Image Understanding
Shanji Liu, Kelu Yao, Junxiao Xue +5
General-purpose vision-language models (VLMs) now support strong visual recognition, instruction following, and generation. However, most pretrained visual encoders are built aroun…
cs.CV2025
Falcon: A Remote Sensing Vision-Language Foundation Model (Technical Report)
Kelu Yao, Nuo Xu, Rong Yang +8
This paper introduces a holistic vision-language foundation model tailored for remote sensing, named Falcon. Falcon offers a unified, prompt-based paradigm that effectively execute…
cs.CV2024★ 1 cited
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation
Nuo Xu, Wen Wang, Rong Yang +6
Object-goal navigation is a challenging task that requires guiding an agent to specific objects based on first-person visual observations. The ability of agent to comprehend its su…