3 papers
cs.CV2026
Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models
Sujung Hong, Chanyong Yoon, Seong Jae Hwang
Large diffusion vision-language models (LDVLMs) have recently emerged as a promising alternative to autoregressive models, enabling parallel decoding for efficient inference and le…
cs.CV2025
Contour Information Aware 2D Gaussian Splatting for Image Representation
Masaya Takabe, Hiroshi Watanabe, Sujun Hong +5
Image representation is a fundamental task in computer vision. Recently, Gaussian Splatting has emerged as an efficient representation framework, and its extension to 2D image repr…
cs.CV2024
DragText: Rethinking Text Embedding in Point-based Image Editing
Gayoon Choi, Taejin Jeong, Sujung Hong +1
Point-based image editing enables accurate and flexible control through content dragging. However, the role of text embedding during the editing process has not been thoroughly inv…