3 citations · 3 across the 4 of their papers we have counts for
4 papers
MegActor-: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
Shurong Yang, Huadong Li, Juhao Wu +6
Diffusion models have demonstrated superior performance in the field of portrait animation. However, current approaches relied on either visual or audio modality to control charact…
Towards RGB-NIR Cross-modality Image Registration and Beyond
Huadong Li, Shichao Dong, Jin Wang +5
This paper focuses on the area of RGB(visible)-NIR(near-infrared) cross-modality image registration, which is crucial for many downstream vision tasks to fully leverage the complem…
Implicit Identity Leakage: The Stumbling Block to Improving Deepfake Detection Generalization
Shichao Dong, Jin Wang, Renhe Ji +3
In this paper, we analyse the generalization ability of binary classifiers for the task of deepfake detection. We find that the stumbling block to their generalization is caused by…
Explaining Deepfake Detection by Analysing Image Matching
Shichao Dong, Jin Wang, Jiajun Liang +2
This paper aims to interpret how deepfake detection models learn artifact features of images when just supervised by binary labels. To this end, three hypotheses from the perspecti…