collaborators

5 papers

cs.CV2026

Keyframe-Based Feed-Forward Visual Odometry

Weichen Dai, Wenhan Su, Da Kong +2

The emergence of visual foundation models has revolutionized visual odometry~(VO) and SLAM, enabling pose estimation and dense reconstruction within a single feed-forward network.…

cs.CV2025

CUS-GS: A Compact Unified Structured Gaussian Splatting Framework for Multimodal Scene Representation

Yuhang Ming, Chenxin Fang, Xingyuan Yu +4

Recent advances in Gaussian Splatting based 3D scene representation have shown two major trends: semantics-oriented approaches that focus on high-level understanding but lack expli…

cs.CV2025

3D Scene-Camera Representation with Joint Camera Photometric Optimization

Weichen Dai, Kangcheng Ma, Jiaxin Wang +4

Representing scenes from multi-view images is a crucial task in computer vision with extensive applications. However, inherent photometric distortions in the camera imaging can sig…

cs.RO2025

SLC-SLAM: Semantic-guided Loop Closure using Shared Latent Code for NeRF SLAM

Yuhang Ming, Di Ma, Weichen Dai +4

Targeting the notorious cumulative drift errors in NeRF SLAM, we propose a Semantic-guided Loop Closure using Shared Latent Code, dubbed SLC-SLAM. We argue that latent codes st…

cs.CV2025

VIPeR: Visual Incremental Place Recognition with Adaptive Mining and Continual Learning

Yuhang Ming, Minyang Xu, Xingrui Yang +5

Visual place recognition (VPR) is an essential component of many autonomous and augmented/virtual reality systems. It enables the systems to robustly localize themselves in large-s…