1 paper · 1 filter
Gang Li, Di Xu, Xing Cheng +2
Although vision Transformers have achieved excellent performance as backbone models in many vision tasks, most of them intend to capture global relations of all tokens in an image…