1 paper · 1 filter
Jiashu Liao, Pietro Liò, Marc de Kamps +1
Vision Transformers face a fundamental limitation: standard self-attention jointly processes spatial and channel dimensions, leading to entangled representations that prevent indep…