1 paper · 1 filter
Ke Hao, Yuanzhi Liang, Tingxi Chen +5
Unified multimodal models integrate visual understanding and generation within a single network, yet the two capabilities are commonly optimized as separate tasks. We introduce Gen…