1 paper
Ching-Lin Hsiung, Tian-Sheuan Chang
Current transformer accelerators primarily focus on optimizing self-attention due to its quadratic complexity. However, this focus is less relevant for vision transformers with sho…