5 papers
The Midas Touch for Metric Depth
Yu Ma, Zizhan Guo, Zuyi Xiong +5
Recent advances have markedly improved the cross-scene generalization of relative depth estimation, yet its practical applicability remains limited by the absence of metric scale,…
Two-Stream Interactive Joint Learning of Scene Parsing and Geometric Vision Tasks
Guanfeng Tang, Hongbo Zhao, Ziwei Long +5
Inspired by the human visual system, which operates on two parallel yet interactive streams for contextual and spatial understanding, this article presents Two Interactive Streams…
Environment-Driven Online LiDAR-Camera Extrinsic Calibration
Zhiwei Huang, Jiaqi Li, Hongbo Zhao +5
LiDAR-camera extrinsic calibration (LCEC) is crucial for multi-modal data fusion in autonomous robotic systems. Existing methods, whether target-based or target-free, typically rel…
Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning
Mengtan Zhang, Zizhan Guo, Hongbo Zhao +6
Unsupervised learning of depth and ego-motion, two fundamental 3D perception tasks, has made significant strides in recent years. However, most methods treat ego-motion as an auxil…
A Birotation Solution for Relative Pose Problems
Hongbo Zhao, Ziwei Long, Mengtan Zhang +3
Relative pose estimation, a fundamental computer vision problem, has been extensively studied for decades. Existing methods either estimate and decompose the essential matrix or di…