activity
20232025
most citedPrivacy-oriented manipulation of speaker representations

4 citations · 6 across the 7 of their papers we have counts for

collaborators

7 papers

cs.CV2025

MASIV: Toward Material-Agnostic System Identification from Videos

Yizhou Zhao, Haoyu Chen, Chunjiang Liu +7

System identification from videos aims to recover object geometry and governing physical laws. Existing methods integrate differentiable rendering with simulation but rely on prede…

cs.CV2025

Total-Editing: Head Avatar with Editable Appearance, Motion, and Lighting

Yizhou Zhao, Chunjiang Liu, Haoyu Chen +6

Face reenactment and portrait relighting are essential tasks in portrait editing, yet they are typically addressed independently, without much synergy. Most face reenactment method…

cs.SD2024

Revisiting Acoustic Features for Robust ASR

Muhammad A. Shah, Bhiksha Raj

Automatic Speech Recognition (ASR) systems must be robust to the myriad types of noises present in real-world environments including environmental noise, room impulse response, spe…

cs.CV2024

Synergistic Global-space Camera and Human Reconstruction from Videos

Yizhou Zhao, Tuanfeng Y. Wang, Bhiksha Raj +3

Remarkable strides have been made in reconstructing static scenes or human bodies from monocular videos. Yet, the two problems have largely been approached independently, without m…

cs.LG2024

Improving Membership Inference in ASR Model Auditing with Perturbed Loss Features

Francisco Teixeira, Karla Pizzi, Raphael Olivier +3

Membership Inference (MI) poses a substantial privacy threat to the training data of Automatic Speech Recognition (ASR) systems, while also offering an opportunity to audit these m…

cs.CV2023★ 2 cited

Weakly-Supervised Audio-Visual Segmentation

Shentong Mo, Bhiksha Raj

Audio-visual segmentation is a challenging task that aims to predict pixel-level masks for sound sources in a video. Previous work applied a comprehensive manually designed archite…