3 papers
cs.CV2026
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
Hang Wang, Chao Shen, Chenhao Lin +3
The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video (AIGV) detection methods pri…
cs.CV2025
HCMA: Hierarchical Cross-model Alignment for Grounded Text-to-Image Generation
Hang Wang, Zhi-Qi Cheng, Chenhao Lin +2
Text-to-image synthesis has progressed to the point where models can generate visually compelling images from natural language prompts. Yet, existing methods often fail to reconcil…
cs.CV2024
IVAC-P2L: Leveraging Irregular Repetition Priors for Improving Video Action Counting
Hang Wang, Zhi-Qi Cheng, Youtian Du +1
Video Action Counting (VAC) is crucial in analyzing sports, fitness, and everyday activities by quantifying repetitive actions in videos. However, traditional VAC methods have over…