1 paper
Minsuk Ji, Sanghyeok Lee, Namhyuk Ahn
Despite their impressive realism, modern text-to-image models still struggle with compositionality, often failing to render accurate object counts, attributes, and spatial relation…