17 papers
ID-V2V: Identity-Preserving Video Restylization
Yuancheng Xu, Mingming He, Pablo Salamanca +5
In visual storytelling, human performances are central to creative intent and narrative meaning. However, preserving human identity and performance while enabling flexible visual e…
GPUSimBench: Towards Scalable and Reliable GPU-Accelerated Simulators in Embodied AI
Huzhenyu Zhang, Shenghai Yuan, Wenrui Yan +4
Data-driven embodied AI is rapidly transitioning into a paradigm that scales training through massively parallel simulation, where GPU-accelerated simulators serve as the foundatio…
Go-with-the-Track: Video Compositing and Motion Control with Point Tracking
Koichi Namekata, Yash Kant, Zhizheng Liu +9
Filmmaking demands precise motion control and reference image compositing -- capabilities that existing methods treat separately. Point-track-conditioned image-to-video models rest…
Cinematic Compositing Using Character-Environment-Harmonized Video Generation Models
Tianyi Xiang, Mingming He, Li Ma +1
Cinematic compositing aims to integrate green-screen characters into novel environments while maintaining physical and photometric realism. Previous methods often fail to capture t…
ReAge3D: Re-Aging 3D Faces with View Consistency
Libing Zeng, Li Ma, Mingming He +3
We present a novel framework for realistic and controllable 3D face re-aging which produces highly detailed, identity-preserving results. Existing 3D editing methods, while effecti…
Lighting in Motion: Spatiotemporal HDR Lighting Estimation
Christophe Bolduc, Julien Philip, Li Ma +3
We present Lighting in Motion (LiMo), a diffusion-based approach to spatiotemporal lighting estimation. LiMo targets both realistic high-frequency detail prediction and accurate il…