1 paper
Chunyang Zhang, Zhenhong Sun, Zhicheng Zhang +5
Text-to-image (T2I) generation models often struggle with multi-instance synthesis (MIS), where they must accurately depict multiple distinct instances in a single image based on c…