1 citations · 1 across the 3 of their papers we have counts for
6 papers
Warm Chat: Diffuse Emotion-aware Interactive Talking Head Avatar with Tree-Structured Guidance
Haijie Yang, Zhenyu Zhang, Hao Tang +2
Generative models have advanced rapidly, enabling impressive talking head generation that brings AI to life. However, most existing methods focus solely on one-way portrait animati…
Spatial-Temporal Graph Mamba for Music-Guided Dance Video Synthesis
Hao Tang, Ling Shao, Zhenyu Zhang +2
We propose a novel spatial-temporal graph Mamba (STG-Mamba) for the music-guided dance video synthesis task, i.e., to translate the input music to a dance video. STG-Mamba consists…
Semantic-Guided Diffusion Model for Single-Step Image Super-Resolution
Zihang Liu, Zhenyu Zhang, Hao Tang
Diffusion-based image super-resolution (SR) methods have demonstrated remarkable performance. Recent advancements have introduced deterministic sampling processes that reduce infer…
Follow Your Motion: A Generic Temporal Consistency Portrait Editing Framework with Trajectory Guidance
Haijie Yang, Zhenyu Zhang, Hao Tang +2
Pre-trained conditional diffusion models have demonstrated remarkable potential in image editing. However, they often face challenges with temporal consistency, particularly in the…
ConsistentAvatar: Learning to Diffuse Fully Consistent Talking Head Avatar with Temporal Guidance
Haijie Yang, Zhenyu Zhang, Hao Tang +2
Diffusion models have shown impressive potential on talking head generation. While plausible appearance and talking effect are achieved, these methods still suffer from temporal, 3…
Toward Zero-Shot Learning for Visual Dehazing of Urological Surgical Robots
Renkai Wu, Xianjin Wang, Pengchen Liang +3
Robot-assisted surgery has profoundly influenced current forms of minimally invasive surgery. However, in transurethral suburethral urological surgical robots, they need to work in…