3 papers
cs.CV2026
MASRA: MLLM-Assisted Semantic-Relational Consistent Alignment for Video Temporal Grounding
Ran Ran, Jiwei Wei, Shuchang Zhou +5
Video Temporal Grounding (VTG) faces a cross-modal semantic gap that often leads to background features being incorrectly aligned with the query, while directly matching the query…
cs.CV2026
Enhancing Self-Supervised Talking Head Forgery Detection via a Training-Free Dual-System Framework
Ke Liu, Jiwei Wei, Shuchang Zhou +5
Supervised talking head forgery detection faces severe generalization challenges due to the continuous evolution of generators. By reducing reliance on generator-specific forgery p…
cs.CV2026
HiMix: Hierarchical Artifact-aware Mixup for Generalized Synthetic Image Detection
Shuchang Zhou, Kaiwen Shen, Jiwei Wei +3
The rapid evolution of generative models has enabled the creation of highly realistic and diverse synthetic images, posing significant challenges to reliable and generalizable Synt…