34 citations · 51 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 34 cited
LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models
Yaohui Wang, Xinyuan Chen, Xin Ma +17
This work aims to learn a high-quality text-to-video (T2V) generative model by leveraging a pre-trained text-to-image (T2I) model as a basis. It is a highly desirable yet challengi…
cs.CV2023★ 15 cited
VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking
Limin Wang, Bingkun Huang, Zhiyu Zhao +5
Scale is the primary factor for building a powerful foundation model that could well generalize to a variety of downstream tasks. However, it is still challenging to train video fo…
cs.CV2021★ 2 cited
ForgeryNet -- Face Forgery Analysis Challenge 2021: Methods and Results
Yinan He, Lu Sheng, Jing Shao +19
The rapid progress of photorealistic synthesis techniques has reached a critical point where the boundary between real and manipulated images starts to blur. Recently, a mega-scale…