5 papers
ScaleResfusion: Residual Rectified Flow based on Residual Vector Field
Zhenning Shi, Chen Xu, Junhao Zhang +4
Real-world Image Restoration (Real-IR) aims to recover high-quality (HQ) images from complex and unknown degradations. Recent diffusion-based methods have substantially improved pe…
REVEAL: Reference-Grounded Reasoning for Multimodal Manipulation Detection
Jun Zhou, Bingwen Hu, Yaxiong Wang +4
Multimodal manipulation detection aims to simultaneously identify forged image--text pairs and localize tampered regions, yet existing methods typically rely on memorizing isolated…
SEED: A Benchmark Dataset for Sequential Facial Attribute Editing with Diffusion Models
Yule Zhu, Ping Liu, Zhedong Zheng +1
Diffusion models have recently enabled precise and photorealistic facial editing across a wide range of semantic attributes. Beyond single-step modifications, a growing class of ap…
CLIP-SR: Collaborative Linguistic and Image Processing for Super-Resolution
Bingwen Hu, Heng Liu, Zhedong Zheng +1
Convolutional Neural Networks (CNNs) have significantly advanced Image Super-Resolution (SR), yet most CNN-based methods rely solely on pixel-based transformations, often leading t…
RIGI: Rectifying Image-to-3D Generation Inconsistency via Uncertainty-aware Learning
Jiacheng Wang, Zhedong Zheng, Wei Xu +1
Given a single image of a target object, image-to-3D generation aims to reconstruct its texture and geometric shape. Recent methods often utilize intermediate media, such as multi-…