2 citations · 2 across the 2 of their papers we have counts for
1 paper · 1 filter
Zihao Wang, Yuxiang Wei, Xinpeng Zhou +5
Text-to-image generation has advanced rapidly, yet it still struggles to capture the nuanced user preferences. Existing approaches typically rely on multimodal large language model…