4 papers · 1 filter
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
Rit Gangopadhyay, Jung-Hee Kim, Xien Chen +3
We propose a method to extend foundational monocular depth estimators (FMDEs), trained on perspective images, to fisheye images. Despite being trained on tens of millions of images…
UnCLe: Benchmarking Unsupervised Continual Learning for Depth Completion
Xien Chen, Rit Gangopadhyay, Michael Chu +3
We propose UnCLe, the first standardized benchmark for Unsupervised Continual Learning of a multimodal 3D reconstruction task: Depth completion aims to infer a dense depth map from…
Portrait3D: Text-Guided High-Quality 3D Portrait Generation Using Pyramid Representation and GANs Prior
Yiqian Wu, Hao Xu, Xiangjun Tang +5
Existing neural rendering-based text-to-3D-portrait generation methods typically make use of human geometry prior and diffusion models to obtain guidance. However, relying solely o…
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations
Fengyu Yang, Chao Feng, Ziyang Chen +8
The ability to associate touch with other modalities has huge implications for humans and computational systems. However, multimodal learning with touch remains challenging due to…