38 citations · 41 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 3 cited
Vision Transformer with Quadrangle Attention
Qiming Zhang, Jing Zhang, Yufei Xu +1
Window-based attention has become a popular choice in vision transformers due to its superior performance, lower computational complexity, and less memory footprint. However, the d…
cs.CV2022★ 38 cited
Advancing Plain Vision Transformer Towards Remote Sensing Foundation Model
Di Wang, Qiming Zhang, Yufei Xu +4
Large-scale vision foundation models have made significant progress in visual tasks on natural images, with vision transformers being the primary choice due to their good scalabili…