1 paper
Mengqi Lei, Yihong Wu, Siqi Li +4
Visual recognition relies on understanding the semantics of image tokens and their complex interactions. Mainstream self-attention methods, while effective at modeling global pair-…