collaborators

5 papers

cs.CV2024

MultiColor: Image Colorization by Learning from Multiple Color Spaces

Xiangcheng Du, Zhao Zhou, Yanlong Wang +3

Deep networks have shown impressive performance in the image restoration tasks, such as image colorization. However, we find that previous approaches rely on the digital representa…

cs.CV2024

Cross-Domain Document Layout Analysis Using Document Style Guide

Xingjiao Wu, Luwei Xiao, Xiangcheng Du +5

The document layout analysis (DLA) aims to decompose document images into high-level semantic areas (i.e., figures, tables, texts, and background). Creating a DLA framework with st…

cs.CV2024

Minutes to Seconds: Speeded-up DDPM-based Image Inpainting with Coarse-to-Fine Sampling

Lintao Zhang, Xiangcheng Du, LeoWu TomyEnrique +3

For image inpainting, the existing Denoising Diffusion Probabilistic Model (DDPM) based method i.e. RePaint can produce high-quality images for any inpainting form. It utilizes a p…

cs.CV2024

Fine-Grained Scene Image Classification with Modality-Agnostic Adapter

Yiqun Wang, Zhao Zhou, Xiangcheng Du +3

When dealing with the task of fine-grained scene image classification, most previous works lay much emphasis on global visual features when doing multi-modal feature fusion. In oth…

cs.CV2024

Aggregated Text Transformer for Scene Text Detection

Zhao Zhou, Xiangcheng Du, Yingbin Zheng +1

This paper explores the multi-scale aggregation strategy for scene text detection in natural images. We present the Aggregated Text TRansformer(ATTR), which is designed to represen…