1 citations · 2 across the 9 of their papers we have counts for
1 paper · 1 filter
Jichen Yang, Fangfan Chen, Rohan Kumar Das +2
Traditional vision transformer consists of two parts: transformer encoder and multi-layer perception (MLP). The former plays the role of feature learning to obtain better represent…