5 papers
VoxScene: Anchor-Conditioned Voxel Diffusion for Indoor Scene Arrangement
Haotian Mao, Yuhan Huang, Jiatao Lin +8
We present VoxScene, a novel anchor-conditioned voxel diffusion framework tailored for 3D scene synthesis. Current data-driven layout generation techniques typically rely on boundi…
LIVE-GS: LLM Powers Interactive VR Experience with Physics-Aware Gaussian Splatting
Haotian Mao, Hangyu Zhou, Zhuoxiong Xu +6
As 3D Gaussian Splatting (3DGS) emerges as a leading approach for novel view synthesis and scene reconstruction, its potential in digital asset creation has gained significant atte…
SceneReVis: A Self-Reflective Vision-Grounded Framework for 3D Indoor Scene Synthesis via Multi-turn RL
Yang Zhao, Shizhao Sun, Meisheng Zhang +3
Current one-pass 3D scene synthesis methods often suffer from spatial hallucinations, such as collisions, due to a lack of deliberative reasoning. To bridge this gap, we introduce…
Survey of Large Language Models in Extended Reality: Technical Paradigms and Application Frontiers
Jingyan Wang, Yang Zhao, Haotian Mao +1
Large Language Models (LLMs) have demonstrated remarkable capabilities in natural language understanding and generation, and their integration with Extended Reality (XR) is poised…
IKMo: Image-Keyframed Motion Generation with Trajectory-Pose Conditioned Motion Diffusion Model
Yang Zhao, Yan Zhang, Xubo Yang
Existing human motion generation methods with trajectory and pose inputs operate global processing on both modalities, leading to suboptimal outputs. In this paper, we propose IKMo…