3 papers
cs.CV2026
Towards Active Cross-View Object Geo-Localization
Shunyu Yao, Xiaohan Zhang, Zhuoran Yang +5
Cross-view object geo-localization (CVOGL) typically assumes a fixed query image, overlooking the ability of mobile agents to actively acquire more informative observations. To add…
cs.CV2026
Multi-View Mixture-of-Experts with Vision-Language Reranking for Cross-View Object Geo-Localization
Xuyu Fan, Qi Ming, Zhu Han +6
Cross-view object geo-localization (CVOGL) locates a target in satellite imagery using drone or street-view queries. Existing methods train separate detectors for each viewpoint, l…
cs.CV2026
ChainSpace: A Chained-Reasoning Paradigm for Spatial Intelligence
Xiaohan Zhang, Feng Gu, Xudong Rao +4
Spatial intelligence requires foundation models to maintain coherent spatial state across interactions with the physical world. However, existing data-centric approaches typically…