20 papers
Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation
Haijie Yang, Jindi Bao, Yixuan Dong +6
Audio-driven portrait animation has advanced rapidly with diffusion-based generative models, yet real-time one-shot generation with expressive emotion control remains challenging.…
GaussianEmoTalker: Real-Time Emotional Talking Head Synthesis with Audio-Driven and Blendshape-Based 3D Gaussian Splatting
Haijie Yang, Zhenyu Zhang, Yixuan Dong +2
Audio-driven talking head synthesis has achieved impressive progress in lip synchronization and visual quality, yet generating expressive emotional avatars with controllable intens…
OrthoTryOn: Geometric Orthogonalization for Conflict-Free Unified Fashion Generation
Zhaotong Yang, Ying Tai, Jiahui Zhan +3
Unified fashion generation integrates tasks like virtual try-on and garment reconstruction into a single model to reduce task-specific adaptation costs. However, naive parameter sh…
PhysFlow: Frequency Decoupled with Dual-Field Rectified Flow for Remote Photoplethysmography
Zixu Li, jianjun Qian, Hang Shao +2
Remote Photoplethysmography (rPPG) enables contactless pulse estimation from facial videos, serving as a vital tool for health monitoring. However, current deep learning methods of…
GEM: Generating LiDAR World Model via Deformable Mamba
Yang Wu, Zhaojiang Liu, Qiang Meng +5
World models, which simulate environmental dynamics and generate sensor observations, are gaining increasing attention in autonomous driving. However, progress in LiDAR-based world…
ASGNet: Adaptive Spectrum Guidance Network for Automatic Polyp Segmentation
Yanguang Sun, Hengmin Zhang, Jianjun Qian +2
Early identification and removal of polyps can reduce the risk of developing colorectal cancer. However, the diverse morphologies, complex backgrounds and often concealed nature of…