3 papers
cs.CV2026
Decoupling Semantics and Fingerprints: A Universal Representation for AI-Generated Image Detection
Zhiyuan Wang, Yanxiang Chen, Pengcheng Zhao +2
Detecting AI-generated images across unseen architectures remains challenging, as existing models often overfit to generator-specific fingerprints and semantic content rather than…
cs.CV2024
Audio-Infused Automatic Image Colorization by Exploiting Audio Scene Semantics
Pengcheng Zhao, Yanxiang Chen, Yang Zhao +1
Automatic image colorization is inherently an ill-posed problem with uncertainty, which requires an accurate semantic understanding of scenes to estimate reasonable colors for gray…
cs.CV2024
Multimodal Class-aware Semantic Enhancement Network for Audio-Visual Video Parsing
Pengcheng Zhao, Jinxing Zhou, Yang Zhao +2
The Audio-Visual Video Parsing task aims to recognize and temporally localize all events occurring in either the audio or visual stream, or both. Capturing accurate event semantics…