437 citations · 1.1k across the 28 of their papers we have counts for
10 papers · 1 filter
From Two to One: A New Scene Text Recognizer with Visual Language Modeling Network
Yuxin Wang, Hongtao Xie, Shancheng Fang +3
In this paper, we abandon the dominant complex language model and rethink the linguistic learning process in the scene text recognition. Different from previous methods considering…
Causal Incremental Graph Convolution for Recommender System Retraining
Sihao Ding, Fuli Feng, Xiangnan He +3
Real-world recommender system needs to be regularly retrained to keep with the new data. In this work, we consider how to efficiently retrain graph convolution network (GCN) based…
Cross-Modal Attention Consistency for Video-Audio Unsupervised Learning
Shaobo Min, Qi Dai, Hongtao Xie +3
Cross-modal correlation provides an inherent supervision for video unsupervised representation learning. Existing methods focus on distinguishing different video clips by visual an…
Diverse Part Discovery: Occluded Person Re-identification with Part-Aware Transformer
Yulin Li, Jianfeng He, Tianzhu Zhang +3
Occluded person re-identification (Re-ID) is a challenging task as persons are frequently occluded by various obstacles or other persons, especially in the crowd scenario. To addre…
PERT: A Progressively Region-based Network for Scene Text Removal
Yuxin Wang, Hongtao Xie, Shancheng Fang +2
Scene text removal (STR) contains two processes: text localization and background reconstruction. Through integrating both processes into a single network, previous methods provide…
Causal Intervention for Leveraging Popularity Bias in Recommendation
Yang Zhang, Fuli Feng, Xiangnan He +4
Recommender system usually faces popularity bias issues: from the data perspective, items exhibit uneven (long-tail) distribution on the interaction frequency; from the method pers…