3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2025
D2C: Unlocking the Potential of Continuous Autoregressive Image Generation with Discrete Tokens
Panpan Wang, Liqiang Niu, Fandong Meng +3
In the domain of image generation, latent-based generative models occupy a dominant status; however, these models rely heavily on image tokenizer. To meet modeling requirements, au…
cs.CL2024
Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
Zhibin Lan, Liqiang Niu, Fandong Meng +3
In-image machine translation (IIMT) aims to translate an image containing texts in source language into an image containing translations in target language. In this regard, convent…
cs.CV2023★ 3 cited
WeLayout: WeChat Layout Analysis System for the ICDAR 2023 Competition on Robust Layout Segmentation in Corporate Documents
Mingliang Zhang, Zhen Cao, Juntao Liu +3
In this paper, we introduce WeLayout, a novel system for segmenting the layout of corporate documents, which stands for WeChat Layout Analysis System. Our approach utilizes a sophi…