1 paper
Gaoyang Zhang, Bingtao Fu, Qingnan Fan +5
Text-to-image (T2I) diffusion models excel at generating photorealistic images but often fail to render accurate spatial relationships. We identify two core issues underlying this…