most citedNet: Accurate Panorama Depth Estimation on Spherical Surface

14 citations · 18 across the 7 of their papers we have counts for

collaborators

8 papers

cs.CV2024

MCMat: Multiview-Consistent and Physically Accurate PBR Material Generation

Shenhao Zhu, Lingteng Qiu, Xiaodong Gu +11

Existing 2D methods utilize UNet-based diffusion models to generate multi-view physically-based rendering (PBR) maps but struggle with multi-view inconsistency, while some 3D metho…

cs.CV20241 cited

MVImgNet2.0: A Larger-scale Dataset of Multi-view Images

Xiaoguang Han, Yushuang Wu, Luyue Shi +7

MVImgNet is a large-scale dataset that contains multi-view images of ~220k real-world objects in 238 classes. As a counterpart of ImageNet, it introduces 3D visual signals via mult…

cs.CV2024

HIVE: HIerarchical Volume Encoding for Neural Implicit Surface Reconstruction

Xiaodong Gu, Weihao Yuan, Heng Li +2

Neural implicit surface reconstruction has become a new trend in reconstructing a detailed 3D shape from images. In previous methods, however, the 3D scene is only encoded by the M…

cs.CV2024

Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition

Yisheng He, Weihao Yuan, Siyu Zhu +3

This paper enables high-fidelity, transferable NeRF editing by frequency decomposition. Recent NeRF editing pipelines lift 2D stylization results to 3D scenes while suffering from…

cs.CV20241 cited

OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation

Junhao Cai, Yisheng He, Weihao Yuan +4

This paper studies a new open-set problem, the open-vocabulary category-level object pose and size estimation. Given human text descriptions of arbitrary novel object categories, t…

cs.CV20241 cited

VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model

Qi Zuo, Xiaodong Gu, Lingteng Qiu +8

Generating multi-view images based on text or single-image prompts is a critical capability for the creation of 3D content. Two fundamental questions on this topic are what data we…