1 paper
Xuwei Xu, Yang Li, Yudong Chen +2
We reveal that feedforward network (FFN) layers, rather than attention layers, are the primary contributors to Vision Transformer (ViT) inference latency, with their impact signify…