4 papers · 1 filter
Let Triggers Control: Frequency-Aware Dropout for Effective Token Control
Junyoung Koh, Hoyeon Moon, Dongha Kim +3
Text-to-image models such as Stable Diffusion have achieved unprecedented levels of high-fidelity visual synthesis. As these models advance, personalization of generative models --…
Improving Text Generation on Images with Synthetic Captions
Jun Young Koh, Sang Hyun Park, Joy Song
The recent emergence of latent diffusion models such as SDXL and SD 1.5 has shown significant capability in generating highly detailed and realistic images. Despite their remarkabl…
CAT: Contrastive Adapter Training for Personalized Image Generation
Jae Wan Park, Sang Hyun Park, Jun Young Koh +2
The emergence of various adapters, including Low-Rank Adaptation (LoRA) applied from the field of natural language processing, has allowed diffusion models to personalize image gen…
Illustrious: an Open Advanced Illustration Model
Sang Hyun Park, Jun Young Koh, Junha Lee +5
In this work, we share the insights for achieving state-of-the-art quality in our text-to-image anime image generative model, called Illustrious. To achieve high resolution, dynami…