3 papers
cs.IR2024
Leveraging Large Vision-Language Model as User Intent-aware Encoder for Composed Image Retrieval
Zelong Sun, Dong Jing, Guoxing Yang +2
Composed Image Retrieval (CIR) aims to retrieve target images from candidate set using a hybrid-modality query consisting of a reference image and a relative caption that describes…
cs.CV2024
Artifact Feature Purification for Cross-domain Detection of AI-generated Images
Zheling Meng, Bo Peng, Jing Dong +1
In the era of AIGC, the fast development of visual content generation technologies, such as diffusion models, bring potential security risks to our society. Existing generated imag…
cs.CV2023
GaFET: Learning Geometry-aware Facial Expression Translation from In-The-Wild Images
Tianxiang Ma, Bingchuan Li, Qian He +2
While current face animation methods can manipulate expressions individually, they suffer from several limitations. The expressions manipulated by some motion-based facial reenactm…