3 citations · 4 across the 25 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Visual Autoregressive Modeling for Instruction-Guided Image Editing
Qingyang Mao, Qi Cai, Yehao Li +5
Recent advances in diffusion models have brought remarkable visual fidelity to instruction-guided image editing. However, their global denoising process inherently entangles the ed…
cs.CV2025★ 1 cited
TextMatch: Enhancing Image-Text Consistency Through Multimodal Optimization
Yucong Luo, Mingyue Cheng, Jie Ouyang +2
Text-to-image generative models excel in creating images from text but struggle with ensuring alignment and consistency between outputs and prompts. This paper introduces TextMatch…