15 citations · 15 across the 1 of their papers we have counts for
1 paper
Yuanfeng Ji, Ruimao Zhang, Huijie Wang +4
The recent vision transformer(i.e.for image classification) learns non-local attentive interaction of different patch tokens. However, prior arts miss learning the cross-scale depe…