1 citations · 1 across the 1 of their papers we have counts for
1 paper
Jichen Yang, Fangfan Chen, Rohan Kumar Das +2
Traditional vision transformer consists of two parts: transformer encoder and multi-layer perception (MLP). The former plays the role of feature learning to obtain better represent…