14 citations · 18 across the 7 of their papers we have counts for
8 papers
MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation
Shenhao Zhu, Lingteng Qiu, Xiaodong Gu +11
Existing 2D methods utilize UNet-based diffusion models to generate multi-view physically-based rendering (PBR) maps but struggle with multi-view inconsistency, while some 3D metho…
MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
Xiaoguang Han, Yushuang Wu, Luyue Shi +7
MVImgNet is a large-scale dataset that contains multi-view images of ~220k real-world objects in 238 classes. As a counterpart of ImageNet, it introduces 3D visual signals via mult…
HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction
Xiaodong Gu, Weihao Yuan, Heng Li +2
Neural implicit surface reconstruction has become a new trend in reconstructing a detailed 3D shape from images. In previous methods, however, the 3D scene is only encoded by the M…
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
Yisheng He, Weihao Yuan, Siyu Zhu +3
This paper enables high-fidelity, transferable NeRF editing by frequency decomposition. Recent NeRF editing pipelines lift 2D stylization results to 3D scenes while suffering from…
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
Junhao Cai, Yisheng He, Weihao Yuan +4
This paper studies a new open-set problem, the open-vocabulary category-level object pose and size estimation. Given human text descriptions of arbitrary novel object categories, t…
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
Qi Zuo, Xiaodong Gu, Lingteng Qiu +8
Generating multi-view images based on text or single-image prompts is a critical capability for the creation of 3D content. Two fundamental questions on this topic are what data we…