5 citations · 5 across the 3 of their papers we have counts for
1 paper · 1 filter
Siheng Wan, Zhengtao Yao, Zhengdao Li +9
Modern Text-to-Image (T2I) generation increasingly relies on token-centric architectures that are trained with self-supervision, yet effectively fusing text with visual tokens rema…