activity
20212024
most citedCAPro: Webly Supervised Learning with Cross-Modality Aligned Prototypes

1 citations · 5 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV20231 cited

Looking and Listening: Audio Guided Text Recognition

Wenwen Yu, Mingyu Liu, Biao Yang +5

Text recognition in the wild is a long-standing problem in computer vision. Driven by end-to-end deep learning, recent studies suggest vision and language processing are effective…

cs.CV2023

ICDAR 2023 Competition on Structured Text Extraction from Visually-Rich Document Images

Wenwen Yu, Chengquan Zhang, Haoyu Cao +24

Structured text extraction is one of the most valuable and challenging application directions in the field of Document AI. However, the scenarios of past benchmarks are limited, an…

cs.CV20231 cited

Grab What You Need: Rethinking Complex Table Structure Recognition with Flexible Components Deliberation

Hao Liu, Xin Li, Mingming Gong +5

Recently, Table Structure Recognition (TSR) task, aiming at identifying table structure into machine readable formats, has received increasing interest in the community. While impr…

cs.CV2023

Co-Salient Object Detection with Co-Representation Purification

Ziyue Zhu, Zhao Zhang, Zheng Lin +2

Co-salient object detection (Co-SOD) aims at discovering the common objects in a group of relevant images. Mining a co-representation is essential for locating co-salient objects.…

cs.CV20211 cited

Reciprocal Normalization for Domain Adaptation

Zhiyong Huang, Kekai Sheng, Ke Li +5

Batch normalization (BN) is widely used in modern deep neural networks, which has been shown to represent the domain-related knowledge, and thus is ineffective for cross-domain tas…