activity
20242026
collaborators

6 papers

cs.CV2026

MAPS: Multi-Anchor Projection Similarity for Joint Vision-Language Geo-Localization

Yutong Hu, Siyuan Tan, Shaocheng Yan +3

Humans localize places by integrating perceptual cues from vision with semantic reasoning from language, forming a scene understanding that is both intuitive and structured. Althou…

cs.CV2026

Global Cross-Modal Geo-Localization: A Million-Scale Dataset and a Physical Consistency Learning Framework

Yutong Hu, Jinhui Chen, Chaoqiang Xu +6

Cross-modal Geo-localization (CMGL) matches ground-level text descriptions with geo-tagged aerial imagery, which is crucial for pedestrian navigation and emergency response. Howeve…

cs.CV2026

ULF-Loc: Unbiased Landmark Feature for Robust Visual Localization with 3D Gaussian Splatting

Yingdong Gu, Shaocheng Yan, Zhenjun Zhao +4

Visual localization is a core technology for augmented reality and autonomous navigation. Recent methods combine the efficient rendering of 3D Gaussian Splatting (3DGS) with featur…

cs.CV2026

Advances in Global Solvers for 3D Vision

Zhenjun Zhao, Heng Yang, Bangyan Liao +7

Global solvers have emerged as a powerful paradigm for 3D vision, offering certifiable solutions to nonconvex geometric optimization problems traditionally addressed by local or he…

cs.CV2025

TurboReg: TurboClique for Robust and Efficient Point Cloud Registration

Shaocheng Yan, Pengcheng Shi, Zhenjun Zhao +4

Robust estimation is essential in correspondence-based Point Cloud Registration (PCR). Existing methods using maximal clique search in compatibility graphs achieve high recall but…

cs.CV2024

RANSAC Back to SOTA: A Two-stage Consensus Filtering for Real-time 3D Registration

Pengcheng Shi, Shaocheng Yan, Yilin Xiao +3

Correspondence-based point cloud registration (PCR) plays a key role in robotics and computer vision. However, challenges like sensor noises, object occlusions, and descriptor limi…