1 paper
Peng Sun, Jun Xie, Tao Lin
Unified Multimodal Models (UMMs) are often constrained by the pre-training of their visual generation components, which typically relies on inefficient paradigms and sca…