61 citations · 117 across the 6 of their papers we have counts for
6 papers
MegActor-: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
Shurong Yang, Huadong Li, Juhao Wu +6
Diffusion models have demonstrated superior performance in the field of portrait animation. However, current approaches relied on either visual or audio modality to control charact…
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
Shurong Yang, Huadong Li, Juhao Wu +5
Despite raw driving videos contain richer information on facial expressions than intermediate representations such as landmarks in the field of portrait animation, they are seldom…
GAFlow: Incorporating Gaussian Attention into Optical Flow
Ao Luo, Fan Yang, Xin Li +4
Optical flow, or the estimation of motion fields from image sequences, is one of the fundamental problems in computer vision. Unlike most pixel-wise tasks that aim at achieving con…
Supervised Homography Learning with Realistic Dataset Generation
Hai Jiang, Haipeng Li, Songchen Han +3
In this paper, we propose an iterative framework, which consists of two phases: a generation phase and a training phase, to generate realistic training data and yield a supervised…
A Point Set Generation Network for 3D Object Reconstruction from a Single Image
Haoqiang Fan, Hao Su, Leonidas Guibas
Generation of 3D data by deep neural network has been attracting increasing attention in the research community. The majority of extant works resort to regular representations such…
Learning Deep Face Representation
Haoqiang Fan, Zhimin Cao, Yuning Jiang +2
Face representation is a crucial step of face recognition systems. An optimal face representation should be discriminative, robust, compact, and very easy-to-implement. While numer…