From the 1 of 4 linked papers with an AI index.
4 papers
Towards Consistent Video Geometry Estimation
Zhu Yu, Jingnan Gao, Runmin Zhang +9
ViGeo is a transformer-based model that estimates dense, temporally consistent geometry (depth, surface normals, and point maps) from video sequences using dynamic chunking attenti…
Large Depth Completion Model from Sparse Observations
Zhu Yu, Zhengyi Zhao, Runmin Zhang +7
This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with sparse observations. Without rely…
LAM: Large Avatar Model for One-shot Animatable Gaussian Head
Yisheng He, Xiaodong Gu, Xiaodan Ye +6
We present LAM, an innovative Large Avatar Model for animatable Gaussian head reconstruction from a single image. Unlike previous methods that require extensive training on capture…
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
Shenhao Zhu, Lingteng Qiu, Xiaodong Gu +11
Existing 2D methods utilize UNet-based diffusion models to generate multi-view physically-based rendering (PBR) maps but struggle with multi-view inconsistency, while some 3D metho…