3 papers
cs.CV2025
ACE++: Instruction-Based Image Creation and Editing via Context-Aware Content Filling
Chaojie Mao, Jingfeng Zhang, Yulin Pan +4
We report ACE++, an instruction-based diffusion framework that tackles various image generation and editing tasks. Inspired by the input format for the inpainting task proposed by…
cs.CV2025
INFELM: In-depth Fairness Evaluation of Large Text-To-Image Models
Di Jin, Xing Liu, Yu Liu +7
The rapid development of large language models (LLMs) and large vision models (LVMs) have propelled the evolution of multi-modal AI systems, which have demonstrated the remarkable…
cs.CV2024
ControlEdit: A MultiModal Local Clothing Image Editing Method
Di Cheng, YingJie Shi, ShiXin Sun +3
Multimodal clothing image editing refers to the precise adjustment and modification of clothing images using data such as textual descriptions and visual images as control conditio…