2 papers
cs.CV2025
STELLAR: Scene Text Editor for Low-Resource Languages and Real-World Data
Yongdeuk Seo, Hyun-seok Min, Sungchul Choi
Scene Text Editing (STE) is the task of modifying text content in an image while preserving its visual style, such as font, color, and background. While recent diffusion-based appr…
cs.CV2024
DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scaling
Kyuheon Jung, Yongdeuk Seo, Seongwoo Cho +3
In this paper, we present an effective data augmentation framework leveraging the Large Language Model (LLM) and Diffusion Model (DM) to tackle the challenges inherent in data-scar…