3 papers
cs.CV2025
EasyGenNet: An Efficient Framework for Audio-Driven Gesture Video Generation Based on Diffusion Model
Renda Li, Xiaohua Qi, Qiang Ling +4
Audio-driven cospeech video generation typically involves two stages: speech-to-gesture and gesture-to-video. While significant advances have been made in speech-to-gesture generat…
cs.CV2025
Data-free Knowledge Distillation with Diffusion Models
Xiaohua Qi, Renda Li, Long Peng +6
Recently Data-Free Knowledge Distillation (DFKD) has garnered attention and can transfer knowledge from a teacher neural network to a student neural network without requiring any a…
cs.CV2023
The Second Monocular Depth Estimation Challenge
Jaime Spencer, C. Stella Qian, Michaela Trescakova +40
This paper discusses the results for the second edition of the Monocular Depth Estimation Challenge (MDEC). This edition was open to methods using any form of supervision, includin…