5 papers
LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation
Yukang Chen, Luozhou Wang, Wei Huang +13
We present LongLive-2.0, an NVFP4-based parallel infrastructure throughout the full training and inference workflow of long video generation, addressing speed and memory bottleneck…
InterPhys: Physics-aware Human Motion Synthesis in a Dynamic Scene
Chaoyue Xing, Wei Mao, Miaomiao Liu
This paper tackles the problem of physics-aware human motion synthesis in a dynamic scene. Unlike existing works which mainly tend to generate physically unrealistic motions due to…
Personalizing Causal Audio-Driven Facial Motion via Dynamic Multi-modal Retrieval
Xuangeng Chu, Yu Han, Wei Mao +1
Audio-driven facial animation is essential for immersive digital interaction, yet existing frameworks fail to reconcile real-time streaming with high-fidelity personalization. Curr…
Learning High-Fidelity Cloth Animation via Skinning-Free Image Transfer
Rong Wang, Wei Mao, Changsheng Lu +1
We present a novel method for generating 3D garment deformations from given body poses, which is key to a wide range of applications, including virtual try-on and extended reality.…
BAG: Body-Aligned 3D Wearable Asset Generation
Zhongjin Luo, Yang Li, Mingrui Zhang +8
While recent advancements have shown remarkable progress in general 3D shape generation models, the challenge of leveraging these approaches to automatically generate wearable 3D a…