activity
20182023
most citedBi-directional Cross-Modality Feature Propagation with Separation-and-Aggregation Gate for RGB-D Semantic Segmentation

51 citations · 306 across the 39 of their papers we have counts for

collaborators
Showing 2021 · cs.CVShow all

7 papers · 2 filters

cs.CV2021

MoCaNet: Motion Retargeting in-the-wild via Canonicalization Networks

Wentao Zhu, Zhuoqian Yang, Ziang Di +3

We present a novel framework that brings the 3D motion retargeting task from controlled environments to in-the-wild scenarios. In particular, our method is capable of retargeting b…

cs.CV2021★ 1 cited

Progressive Attention on Multi-Level Dense Difference Maps for Generic Event Boundary Detection

Jiaqi Tang, Zhaoyang Liu, Chen Qian +2

Generic event boundary detection is an important yet challenging task in video understanding, which aims at detecting the moments where humans naturally perceive event boundaries.…

cs.CV2021

Deceive D: Adaptive Pseudo Augmentation for GAN Training with Limited Data

Liming Jiang, Bo Dai, Wayne Wu +1

Generative adversarial networks (GANs) typically require ample data for training in order to synthesize high-fidelity images. Recent studies have shown that training GANs with limi…

cs.CV2021★ 9 cited

Audio-Driven Emotional Video Portraits

Xinya Ji, Hang Zhou, Kaisiyuan Wang +4

Despite previous success in generating audio-driven talking heads, most of the previous studies focus on the correlation between speech content and the mouth shape. Facial emotion,…

cs.CV2021★ 24 cited

Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation

Hang Zhou, Yasheng Sun, Wayne Wu +3

While accurate lip synchronization has been achieved for arbitrary-subject audio-driven talking face generation, the problem of how to efficiently drive the head pose remains. Prev…

cs.CV2021★ 2 cited

Everything's Talkin': Pareidolia Face Reenactment

Linsen Song, Wayne Wu, Chaoyou Fu +3

We present a new application direction named Pareidolia Face Reenactment, which is defined as animating a static illusory face to move in tandem with a human face in the video. For…