4 papers
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
Hang Wang, Chao Shen, Chenhao Lin +3
The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video (AIGV) detection methods pri…
ATSS: Detecting AI-Generated Videos via Anomalous Temporal Self-Similarity
Hang Wang, Chao Shen, Lei Zhang +1
AI-generated videos (AIGVs) have achieved unprecedented photorealism, posing severe threats to digital forensics. Existing AIGV detectors focus mainly on localized artifacts or sho…
Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis
Zhoulin Ji, Chenhao Lin, Hang Wang +1
Detecting synthetic from real speech is increasingly crucial due to the risks of misinformation and identity impersonation. While various datasets for synthetic speech analysis hav…
HCMA: Hierarchical Cross-model Alignment for Grounded Text-to-Image Generation
Hang Wang, Zhi-Qi Cheng, Chenhao Lin +2
Text-to-image synthesis has progressed to the point where models can generate visually compelling images from natural language prompts. Yet, existing methods often fail to reconcil…