2 papers
cs.CV2026
Semantic One-Dimensional Tokenizer for Image Reconstruction and Generation
Yunpeng Qu, Kaidong Zhang, Yukang Ding +2
Visual generative models based on latent space have achieved great success, underscoring the significance of visual tokenization. Mapping images to latents boosts efficiency and en…
cs.CV2024
Fine-grained Text to Image Synthesis
Xu Ouyang, Ying Chen, Kaiyue Zhu +1
Fine-grained text to image synthesis involves generating images from texts that belong to different categories. In contrast to general text to image synthesis, in fine-grained synt…