4 papers
LPA3D: 3D Room-Level Scene Generation from In-the-Wild Images
Ming-Jia Yang, Yu-Xiao Guo, Yang Liu +2
Generating realistic, room-level indoor scenes with semantically plausible and detailed appearances from in-the-wild images is crucial for various applications in VR, AR, and robot…
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
Sicheng Xu, Guojun Chen, Yu-Xiao Guo +6
We introduce VASA, a framework for generating lifelike talking faces with appealing visual affective skills (VAS) given a single static image and a speech audio clip. Our premiere…
MVD: Efficient Multiview 3D Reconstruction for Multiview Diffusion
Xin-Yang Zheng, Hao Pan, Yu-Xiao Guo +2
As a promising 3D generation technique, multiview diffusion (MVD) has received a lot of attention due to its advantages in terms of generalizability, quality, and efficiency. By fi…
Swin3D++: Effective Multi-Source Pretraining for 3D Indoor Scene Understanding
Yu-Qi Yang, Yu-Xiao Guo, Yang Liu
Data diversity and abundance are essential for improving the performance and generalization of models in natural language processing and 2D vision. However, 3D vision domain suffer…