1 citations · 1 across the 1 of their papers we have counts for
4 papers · 1 filter
Cross-modulated Attention Transformer for RGBT Tracking
Yun Xiao, Jiacong Zhao, Andong Lu +4
Existing Transformer-based RGBT trackers achieve remarkable performance benefits by leveraging self-attention to extract uni-modal features and cross-attention to enhance multi-mod…
Exploring Part-Informed Visual-Language Learning for Person Re-Identification
Yin Lin, Yehansen Chen, Baocai Yin +4
Recently, visual-language learning (VLL) has shown great potential in enhancing visual-based person re-identification (ReID). Existing VLL-based ReID methods typically focus on ima…
Weakly Supervised Scene Text Generation for Low-resource Languages
Yangchen Xie, Xinyuan Chen, Hongjian Zhan +4
A large number of annotated training images is crucial for training successful scene text recognition models. However, collecting sufficient datasets can be a labor-intensive and c…
VGTS: Visually Guided Text Spotting for Novel Categories in Historical Manuscripts
Wenbo Hu, Hongjian Zhan, Xinchen Ma +3
In the field of historical manuscript research, scholars frequently encounter novel symbols in ancient texts, investing considerable effort in their identification and documentatio…