collaborators

5 papers

cs.CV2025

TADoc: Robust Time-Aware Document Image Dewarping

Fangmin Zhao, Weichao Zeng, Zhenhang Li +2

Flattening curved, wrinkled, and rotated document images captured by portable photographing devices, termed document image dewarping, has become an increasingly important task with…

cs.CV2025

Uni-DocDiff: A Unified Document Restoration Model Based on Diffusion

Fangmin Zhao, Weichao Zeng, Zhenhang Li +4

Removing various degradations from damaged documents greatly benefits digitization, downstream document analysis, and readability. Previous methods often treat each restoration tas…

cs.CV2025

The Role of Video Generation in Enhancing Data-Limited Action Understanding

Wei Li, Dezhao Luo, Dongbao Yang +3

Video action understanding tasks in real-world scenarios always suffer data limitations. In this paper, we address the data-limited action understanding problem by bridging data sc…

cs.CV2025

Visual Text Processing: A Comprehensive Review and Unified Evaluation

Yan Shu, Weichao Zeng, Fangmin Zhao +9

Visual text is a crucial component in both document and scene images, conveying rich semantic information and attracting significant attention in the computer vision community. Bey…

cs.CV2025

Beyond Flat Text: Dual Self-inherited Guidance for Visual Text Generation

Minxing Luo, Zixun Xia, Liaojun Chen +7

In real-world images, slanted or curved texts, especially those on cans, banners, or badges, appear as frequently, if not more so, than flat texts due to artistic design or layout…