3 papers
cs.CV2026
VIMCAN: Visual-Inertial 3D Human Pose Estimation with Hybrid Mamba-Cross-Attention Network
Zepeng Yang, Junxuan Bai, Hao Li +4
The rapid advances in deep learning have significantly enhanced the accuracy of multimodal 3D human pose estimation (HPE). However, the state-of-the-art (SOTA) HPE pipelines still…
cs.CV2025
RiemanLine: Riemannian Manifold Representation of 3D Lines for Factor Graph Optimization
Yan Li, Ze Yang, Keisuke Tateno +3
Minimal parametrization of 3D lines plays a critical role in camera localization and structural mapping. Existing representations in robotics and computer vision predominantly hand…
cs.RO2024
Open-Structure: Structural Benchmark Dataset for SLAM Algorithms
Yanyan Li, Zhao Guo, Ze Yang +3
This paper presents Open-Structure, a novel benchmark dataset for evaluating visual odometry and SLAM methods. Compared to existing public datasets that primarily offer raw images,…