activity
20172022
most citedREDBEE: A Visual-Inertial Drone System for Real-Time Moving Object Detection

8 citations · 13 across the 5 of their papers we have counts for

collaborators
Showing cs.CVShow all

7 papers · 1 filter

cs.CV2022

VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild

Kun Cheng, Xiaodong Cun, Yong Zhang +6

We present VideoReTalking, a new system to edit the faces of a real-world talking head video according to input audio, producing a high-quality and lip-syncing output video even wi…

cs.CV2019

Facial age estimation by deep residual decision making

Shichao Li, Kwang-Ting Cheng

Residual representation learning simplifies the optimization problem of learning complex functions and has been widely used by traditional convolutional neural networks. However, i…

cs.CV2019

Multi-Frame Content Integration with a Spatio-Temporal Attention Mechanism for Person Video Motion Transfer

Kun Cheng, Hao-Zhi Huang, Chun Yuan +2

Existing person video generation methods either lack the flexibility in controlling both the appearance and motion, or fail to preserve detailed appearance and temporal consistency…

cs.CV20194 cited

Visualizing the decision-making process in deep neural decision forest

Shichao Li, Kwang-Ting Cheng

Deep neural decision forest (NDF) achieved remarkable performance on various vision tasks via combining decision tree and deep representation learning. In this work, we first trace…

cs.CV2019

MetaPruning: Meta Learning for Automatic Neural Network Channel Pruning

Zechun Liu, Haoyuan Mu, Xiangyu Zhang +4

In this paper, we propose a novel meta learning approach for automatic channel pruning of very deep neural networks. We first train a PruningNet, a kind of meta network, which is a…

cs.CV2018

Bi-Real Net: Binarizing Deep Network Towards Real-Network Performance

Zechun Liu, Wenhan Luo, Baoyuan Wu +3

In this paper, we study 1-bit convolutional neural networks (CNNs), of which both the weights and activations are binary. While efficient, the lacking of representational capabilit…