3 papers
cs.CV2026
EPIC: Efficient Predicate-Guided Inference-Time Control for Compositional Text-to-Image Generation
Sunung Mun, Sunghyun Cho, Jungseul Ok
Recent text-to-image (T2I) generators can synthesize realistic images, but still struggle with compositional prompts involving multiple objects, counts, attributes, and relations.…
cs.CV2026
Edge-Aware Image Manipulation via Diffusion Models with a Novel Structure-Preservation Loss
Minsu Gong, Nuri Ryu, Jungseul Ok +1
Recent advances in image editing leverage latent diffusion models (LDMs) for versatile, text-prompt-driven edits across diverse tasks. Yet, maintaining pixel-level edge structures-…
cs.CV2025
Addressing Text Embedding Leakage in Diffusion-based Image Editing
Sunung Mun, Jinhwan Nam, Sunghyun Cho +1
Text-based image editing, powered by generative diffusion models, lets users modify images through natural-language prompts and has dramatically simplified traditional workflows. D…