1 paper
Yinuo Liu, Zi Qian, Heng Zhou +7
Interleaved text-and-image generation represents a significant frontier for Multimodal Large Language Models (MLLMs), offering a more intuitive way to convey complex information. C…