activity
20242026
collaborators

10 papers

cs.CV2026

Video Understanding by Design: How Datasets Shape Video Models

Lei Wang, Syuan-Hao Li, Piotr Koniusz +1

Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While existing surveys typically organize progr…

cs.GR2026

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation

Zhicheng Zhang, Lei Wang, Yu Zhang +1

Audio-driven talking-head generation has advanced rapidly, yet existing evaluation protocols mainly rely on frame-wise metrics that assume strict temporal correspondence between ge…

cs.CV2026

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation

Zhicheng Zhang, Lei Wang, Yu Zhang +1

Audio-driven talking-head generation has achieved remarkable progress with recent models such as AniTalker, FLOAT, and Sonic. Despite their success, most existing approaches rely o…

cs.CV2026

Trust-Aware Joint Feature-Prediction Discrepancy for Robust Domain Adaptation

Xi Ding, Lei Wang, Syuan-Hao Li +1

Domain adaptation aims to mitigate performance degradation caused by distribution shifts between a labeled source domain and an unlabeled or sparsely labeled target domain. Most ex…

cs.CV2026

Uncertainty-DTW for Sequences and Visual Tokens

Lei Wang, Syuan-Hao Li, Yongsheng Gao +1

Aligning structured data is a fundamental problem in computer vision and machine learning, underlying tasks such as time series analysis, human action recognition, and visual repre…

cs.CV2026

Privacy-Aware Video Anomaly Detection through Orthogonal Subspace Projection

Lei Wang, Wenxiang Diao, Andrew Busch +2

Video anomaly detection (VAD) systems often prioritize accuracy while overlooking privacy concerns, limiting their suitability for real-world deployment. We propose the Orthogonal…