10 papers
Video Understanding by Design: How Datasets Shape Video Models
Lei Wang, Syuan-Hao Li, Piotr Koniusz +1
Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While existing surveys typically organize progr…
Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation
Zhicheng Zhang, Lei Wang, Yu Zhang +1
Audio-driven talking-head generation has advanced rapidly, yet existing evaluation protocols mainly rely on frame-wise metrics that assume strict temporal correspondence between ge…
Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation
Zhicheng Zhang, Lei Wang, Yu Zhang +1
Audio-driven talking-head generation has achieved remarkable progress with recent models such as AniTalker, FLOAT, and Sonic. Despite their success, most existing approaches rely o…
Trust-Aware Joint Feature-Prediction Discrepancy for Robust Domain Adaptation
Xi Ding, Lei Wang, Syuan-Hao Li +1
Domain adaptation aims to mitigate performance degradation caused by distribution shifts between a labeled source domain and an unlabeled or sparsely labeled target domain. Most ex…
Uncertainty-DTW for Sequences and Visual Tokens
Lei Wang, Syuan-Hao Li, Yongsheng Gao +1
Aligning structured data is a fundamental problem in computer vision and machine learning, underlying tasks such as time series analysis, human action recognition, and visual repre…
Privacy-Aware Video Anomaly Detection through Orthogonal Subspace Projection
Lei Wang, Wenxiang Diao, Andrew Busch +2
Video anomaly detection (VAD) systems often prioritize accuracy while overlooking privacy concerns, limiting their suitability for real-world deployment. We propose the Orthogonal…