11 citations · 11 across the 1 of their papers we have counts for
1 paper · 1 filter
Lin Song, Songyang Zhang, Songtao Liu +5
Transformers, the de-facto standard for language modeling, have been recently applied for vision tasks. This paper introduces sparse queries for vision transformers to exploit the…