2 citations · 2 across the 1 of their papers we have counts for
1 paper
Yichen Zhu, Yuqin Zhu, Jie Du +4
The vision transformer splits each image into a sequence of tokens with fixed length and processes the tokens in the same way as words in natural language processing. More tokens n…