2 papers
cs.LG2026
Training Crossroads for Recurrent Vision Transformers: Recurrence, Neural ODEs, and Deep Supervision
Grzegorz Gruszczynski, Pawel Olszowiec, Michal Byra +2
Vision Transformers (ViTs) achieve strong image-recognition performance, but their parameter count grows linearly with depth when each block is independently parameterized. Single-…
cs.CV2026
bViT: Investigating Single-Block Recurrence in Vision Transformers for Image Recognition
Michal Byra, Pawel Olszowiec, Grzegorz Stefanski +2
Vision Transformers (ViTs) are built by stacking independently parameterized blocks, but it remains unclear how much of this depth requires layer specific transformations and how m…