1 paper · 1 filter
Ethan Knights
For state-of-the-art image understanding, Vision Transformers (ViTs) have become the standard architecture but their processing diverges substantially from human attentional charac…