89 citations · 255 across the 14 of their papers we have counts for
14 papers
Freditor: High-Fidelity and Transferable NeRF Editing by Frequency Decomposition
Yisheng He, Weihao Yuan, Siyu Zhu +3
This paper enables high-fidelity, transferable NeRF editing by frequency decomposition. Recent NeRF editing pipelines lift 2D stylization results to 3D scenes while suffering from…
IPoD: Implicit Field Learning with Point Diffusion for Generalizable 3D Object Reconstruction from Single RGB-D Images
Yushuang Wu, Luyue Shi, Junhao Cai +6
Generalizable 3D object reconstruction from single-view RGB-D images remains a challenging task, particularly with real-world data. Current state-of-the-art methods develop Transfo…
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
Junhao Cai, Yisheng He, Weihao Yuan +4
This paper studies a new open-set problem, the open-vocabulary category-level object pose and size estimation. Given human text descriptions of arbitrary novel object categories, t…
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
Qi Zuo, Xiaodong Gu, Lingteng Qiu +8
Generating multi-view images based on text or single-image prompts is a critical capability for the creation of 3D content. Two fundamental questions on this topic are what data we…
Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation
Minglin Chen, Weihao Yuan, Yukun Wang +5
Recently, text-to-3D approaches have achieved high-fidelity 3D content generation using text description. However, the generated objects are stochastic and lack fine-grained contro…
DanceMeld: Unraveling Dance Phrases with Hierarchical Latent Codes for Music-to-Dance Synthesis
Xin Gao, Li Hu, Peng Zhang +2
In the realm of 3D digital human applications, music-to-dance presents a challenging task. Given the one-to-many relationship between music and dance, previous methods have been li…