5 papers
WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes
Jichen Hu, Jiawei Guo, Jiazhong Cen +3
Recent 3D world modeling systems based on generative scene synthesis, such as Marble, can create coherent and explorable 3D environments, yet their outputs are typically static mon…
UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars
Xiaoyu Zhan, Xinyu Fu, Chenghao Yang +9
Speech-driven gestures and facial animations are fundamental to expressive digital avatars in games, virtual production, and interactive media. However, existing methods are either…
Dereflection Any Image with Diffusion Priors and Diversified Data
Jichen Hu, Chen Yang, Zanwei Zhou +4
Reflection removal of a single image remains a highly challenging task due to the complex entanglement between target scenes and unwanted reflections. Despite significant progress,…
Segment Any 3D Gaussians
Jiazhong Cen, Jiemin Fang, Chen Yang +4
This paper presents SAGA (Segment Any 3D GAussians), a highly efficient 3D promptable segmentation method based on 3D Gaussian Splatting (3D-GS). Given 2D visual prompts as input,…
Realistic Surgical Simulation from Monocular Videos
Kailing Wang, Chen Yang, Keyang Zhao +2
This paper tackles the challenge of automatically performing realistic surgical simulations from readily available surgical videos. Recent efforts have successfully integrated phys…