1 paper · 1 filter
Kyumin Choi, Ikbeom Jang
Vision Transformers (ViTs) achieve strong performance but suffer from high computational costs due to quadratic self-attention complexity. Although token reduction techniques such…