1 paper
Jungmyung Wi, Hyunsoo Kim, Donghyun Kim
Text-to-image models produce images that align well with natural language prompts, but compositional generation has long been a central challenge. Models often struggle to satisfy…