1 paper
Xiaoxiao Ma, Mohan Zhou, Tao Liang +5
We introduce STAR, a text-to-image model that employs a scale-wise auto-regressive paradigm. Unlike VAR, which is constrained to class-conditioned synthesis for images up to 256$\t…