activity
20162026
most citedEmotional Speech-Driven Animation with Content-Emotion Disentanglement

82 citations · 144 across the 13 of their papers we have counts for

collaborators
Showing cs.CVShow all

16 papers · 1 filter

cs.CV2026

Streaming Autoregressive Video Generation via Diagonal Distillation

Jinxiu Liu, Xuanming Liu, Kangfu Mei +3

Large pretrained diffusion models have significantly enhanced the quality of generated videos, and yet their use in real-time streaming remains limited. Autoregressive models offer…

cs.CV2024

Towards Variable and Coordinated Holistic Co-Speech Motion Generation

Yifei Liu, Qiong Cao, Yandong Wen +2

This paper addresses the problem of generating lifelike holistic co-speech motions for 3D avatars, focusing on two key aspects: variability and coordination. Variability allows the…

cs.CV2023★ 1 cited

Pairwise Similarity Learning is SimPLE

Yandong Wen, Weiyang Liu, Yao Feng +5

In this paper, we focus on a general yet important learning problem, pairwise similarity learning (PSL). PSL subsumes a wide range of important applications, such as open-set face…

cs.CV2023★ 1 cited

Text-Guided Generation and Editing of Compositional 3D Avatars

Hao Zhang, Yao Feng, Peter Kulits +3

Our goal is to create a realistic 3D facial avatar with hair and accessories using only a text description. While this challenge has attracted significant recent interest, existing…

cs.CV2023

The Hidden Dance of Phonemes and Visage: Unveiling the Enigmatic Link between Phonemes and Facial Features

Liao Qu, Xianwei Zou, Xiang Li +3

This work unveils the enigmatic link between phonemes and facial features. Traditional studies on voice-face correlations typically involve using a long period of voice input, incl…

cs.CV2023

Rethinking Voice-Face Correlation: A Geometry View

Xiang Li, Yandong Wen, Muqiao Yang +3

Previous works on voice-face matching and voice-guided face synthesis demonstrate strong correlations between voice and face, but mainly rely on coarse semantic cues such as gender…