3 papers
cs.CV2026
BindEdit: Taming Attention Leakage for Precise Multi-Object Image Editing
Chaewon Park, Soyoon Lee, Naeun Lee +3
Real image editing enables precise manipulation of visual content, yet existing methods often fail in complex multi-object scenarios, causing semantic blending, object duplication,…
cs.CV2026
Can MLLMs Reason About Visual Persuasion? Evaluating the Efficacy and Faithfulness of Reasoning
Naeun Lee, Hyunjong Kim, Sunghwan Choi +2
Despite strong performance of Multimodal Large Language Models (MLLMs) on multimodal tasks, predicting whether and why an image is persuasive remains challenging. We first show tha…
cs.CV2025
S3D: Sketch-Driven 3D Model Generation
Hail Song, Wonsik Shin, Naeun Lee +3
Generating high-quality 3D models from 2D sketches is a challenging task due to the inherent ambiguity and sparsity of sketch data. In this paper, we present S3D, a novel framework…