13 citations · 18 across the 4 of their papers we have counts for
4 papers · 1 filter
BadWorld: Adversarial Attacks on World Models
Linghui Shen, Mingyue Cui, Xingyi Yang
Visual world models (VWMs) synthesize interactive, action-conditioned rollouts from a single context image. However, it remains an open question how robust these models are to adve…
DaGAN++: Depth-Aware Generative Adversarial Network for Talking Head Video Generation
Fa-Ting Hong, Li Shen, Dan Xu
Predominant techniques on talking head generation largely depend on 2D information, including facial appearances and motions from input face images. Nevertheless, dense 3D facial g…
UniFaceGAN: A Unified Framework for Temporally Consistent Facial Video Editing
Meng Cao, Haozhi Huang, Hao Wang +6
Recent research has witnessed advances in facial image editing tasks including face swapping and face reenactment. However, these methods are confined to dealing with one specific…
Task-agnostic Temporally Consistent Facial Video Editing
Meng Cao, Haozhi Huang, Hao Wang +6
Recent research has witnessed the advances in facial image editing tasks. For video editing, however, previous methods either simply apply transformations frame by frame or utilize…