Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
GrabVG: Graph-Attentive Binding for Visual Grounding in UAV Imagery
Chaowei Wang, Yan Di, Jingjun Sun +5
Visual grounding in Unmanned Aerial Vehicle (UAV) imagery aims to localize a target object in complex bird's-eye-view scenes according to a natural language description. However, t…
cs.CV2026
From Scene-Centric to Observer-Centric: Modeling Observer-Aware Relations for 3D Scene Graph Generation
Jingjun Sun, Chaowei Wang, Zhirui Liu +5
3D Scene Graph Generation (3DSGG) represents 3D scenes as structured object--relation--object graphs for spatial understanding. In observer-centric spatial perception, the same sce…
cs.CV2024
Why mamba is effective? Exploit Linear Transformer-Mamba Network for Multi-Modality Image Fusion
Chenguang Zhu, Shan Gao, Huafeng Chen +5
Multi-modality image fusion aims to integrate the merits of images from different sources and render high-quality fusion images. However, existing feature extraction and fusion met…