1 paper · 1 filter
Jen-Yuan Huang, Tong Lin, Yilun Du
While modern text-to-image (T2I) models excel at generating images from intricate prompts, they struggle to capture the key details when the inputs are descriptive paragraphs. This…