1 paper
Mingcheng Ye, Jiaming Liu, Yiren Song
Interleaved text-image generation aims to jointly produce coherent visual frames and aligned textual descriptions within a single sequence, enabling tasks such as style transfer, c…