1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2024
MegActor-: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
Shurong Yang, Huadong Li, Juhao Wu +6
Diffusion models have demonstrated superior performance in the field of portrait animation. However, current approaches relied on either visual or audio modality to control charact…
cs.CV2024★ 1 cited
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
Shurong Yang, Huadong Li, Juhao Wu +5
Despite raw driving videos contain richer information on facial expressions than intermediate representations such as landmarks in the field of portrait animation, they are seldom…
cs.CV2024
Towards RGB-NIR Cross-modality Image Registration and Beyond
Huadong Li, Shichao Dong, Jin Wang +5
This paper focuses on the area of RGB(visible)-NIR(near-infrared) cross-modality image registration, which is crucial for many downstream vision tasks to fully leverage the complem…