1 citations · 1 across the 1 of their papers we have counts for
3 papers
eess.AS2025
See the Speaker: Crafting High-Resolution Talking Faces from Speech with Prior Guidance and Region Refinement
Jinting Wang, Jun Wang, Hei Victor Cheng +1
Unlike existing methods that rely on source images as appearance references and use source speech to generate motion, this work proposes a novel approach that directly extracts inf…
cs.CV2024★ 1 cited
A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights
Wentao Lei, Jinting Wang, Fengji Ma +2
Human video generation is a dynamic and rapidly evolving task that aims to synthesize 2D human body video sequences with generative models given control conditions such as text, au…
cs.CV2024
TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability
Fengji Ma, Hei Victor Cheng, Chenxing Li +1
Achieving zero-shot adversarial robustness without sacrificing generalization remains challenging for foundation models such as CLIP, especially under large adversarial perturbatio…