1 paper · 1 filter
Zhongbin Guo, Jiahao Xie, Dongling Xiao +5
While Multimodal Large Language Models (MLLMs) have achieved remarkable progress, visual understanding and generation are typically treated as divergent objectives. Existing unifie…