Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Visual Autoregressive Modeling for Instruction-Guided Image Editing
Qingyang Mao, Qi Cai, Yehao Li +5
Recent advances in diffusion models have brought remarkable visual fidelity to instruction-guided image editing. However, their global denoising process inherently entangles the ed…
cs.CV2025
TextMatch: Enhancing Image-Text Consistency Through Multimodal Optimization
Yucong Luo, Mingyue Cheng, Jie Ouyang +2
Text-to-image generative models excel in creating images from text but struggle with ensuring alignment and consistency between outputs and prompts. This paper introduces TextMatch…