1 paper
Qinzhe Yang, Keyan Chen, Jia Xu +2
The computational complexity of Transformers scales quadratically with the number of tokens, which significantly constrains the efficiency of vision models, particularly recent ViT…