6 papers
Line-of-Sight-Constrained Multi-Robot Mapless Navigation via Polygonal Visible Regions
Ruofei Bai, Shenghai Yuan, Xinhang Xu +5
Multi-robot systems rely on underlying connectivity to ensure reliable communication and timely coordination. This paper studies the line-of-sight (LoS) connectivity maintenance pr…
Towards Human-Like Manipulation through RL-Augmented Teleoperation and Mixture-of-Dexterous-Experts VLA
Tutian Tang, Xingyu Ji, Wanli Xing +7
While Vision-Language-Action (VLA) models have demonstrated remarkable success in robotic manipulation, their application has largely been confined to low-degree-of-freedom end-eff…
Stereo-Inertial Poser: Towards Metric-Accurate Shape-Aware Motion Capture Using Sparse IMUs and a Single Stereo Camera
Tutian Tang, Xingyu Ji, Yutong Li +3
Recent advancements in visual-inertial motion capture systems have demonstrated the potential of combining monocular cameras with sparse inertial measurement units (IMUs) as cost-e…
Gaussian Semantic Field for One-shot LiDAR Global Localization
Pengyu Yin, Shenghai Yuan, Haozhi Cao +4
We present a one-shot LiDAR global localization algorithm featuring semantic disambiguation ability based on a lightweight tri-layered scene graph. While landmark semantic registra…
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
Haozhi Cao, Yuecong Xu, Pengyu Yin +4
Multi-modal test-time adaptation (MM-TTA) adapts models to an unlabeled target domain by leveraging the complementary multi-modal inputs in an online manner. While previous MM-TTA…
SGBA: Semantic Gaussian Mixture Model-Based LiDAR Bundle Adjustment
Xingyu Ji, Shenghai Yuan, Jianping Li +3
LiDAR bundle adjustment (BA) is an effective approach to reduce the drifts in pose estimation from the front-end. Existing works on LiDAR BA usually rely on predefined geometric fe…