219 citations · 451 across the 21 of their papers we have counts for
21 papers · 1 filter
Img2Vec: A Teacher of High Token-Diversity Helps Masked AutoEncoders
Heng Pan, Chenyang Liu, Wenxiao Wang +4
We present a pipeline of Image to Vector (Img2Vec) for masked image modeling (MIM) with deep features. To study which type of deep features is appropriate for MIM as a learning tar…
Regional Adversarial Training for Better Robust Generalization
Chuanbiao Song, Yanbo Fan, Yichen Yang +4
Adversarial training (AT) has been demonstrated as one of the most promising defense methods against various adversarial attacks. To our knowledge, existing AT-based methods usuall…
Target Adaptive Context Aggregation for Video Scene Graph Generation
Yao Teng, Limin Wang, Zhifeng Li +1
This paper deals with a challenging task of video scene graph generation (VidSGG), which could serve as a structured video representation for high-level understanding tasks. We pre…
UniFaceGAN: A Unified Framework for Temporally Consistent Facial Video Editing
Meng Cao, Haozhi Huang, Hao Wang +6
Recent research has witnessed advances in facial image editing tasks including face swapping and face reenactment. However, these methods are confined to dealing with one specific…
Structure-Regularized Attention for Deformable Object Representation
Shenao Zhang, Li Shen, Zhifeng Li +1
Capturing contextual dependencies has proven useful to improve the representational power of deep neural networks. Recent approaches that focus on modeling global context, such as…
Image-to-Video Generation via 3D Facial Dynamics
Xiaoguang Tu, Yingtian Zou, Jian Zhao +8
We present a versatile model, FaceAnime, for various video generation tasks from still images. Video generation from a single face image is an interesting problem and usually tackl…