5 papers
MultiColor: Image Colorization by Learning from Multiple Color Spaces
Xiangcheng Du, Zhao Zhou, Yanlong Wang +3
Deep networks have shown impressive performance in the image restoration tasks, such as image colorization. However, we find that previous approaches rely on the digital representa…
Cross-Domain Document Layout Analysis Using Document Style Guide
Xingjiao Wu, Luwei Xiao, Xiangcheng Du +5
The document layout analysis (DLA) aims to decompose document images into high-level semantic areas (i.e., figures, tables, texts, and background). Creating a DLA framework with st…
Minutes to Seconds: Speeded-up DDPM-based Image Inpainting with Coarse-to-Fine Sampling
Lintao Zhang, Xiangcheng Du, LeoWu TomyEnrique +3
For image inpainting, the existing Denoising Diffusion Probabilistic Model (DDPM) based method i.e. RePaint can produce high-quality images for any inpainting form. It utilizes a p…
Fine-Grained Scene Image Classification with Modality-Agnostic Adapter
Yiqun Wang, Zhao Zhou, Xiangcheng Du +3
When dealing with the task of fine-grained scene image classification, most previous works lay much emphasis on global visual features when doing multi-modal feature fusion. In oth…
Aggregated Text Transformer for Scene Text Detection
Zhao Zhou, Xiangcheng Du, Yingbin Zheng +1
This paper explores the multi-scale aggregation strategy for scene text detection in natural images. We present the Aggregated Text TRansformer(ATTR), which is designed to represen…