7 citations · 7 across the 1 of their papers we have counts for
2 papers
cs.CV2022★ 7 cited
Vision Transformers provably learn spatial structure
Samy Jelassi, Michael E. Sander, Yuanzhi Li
Vision Transformers (ViTs) have achieved comparable or superior performance than Convolutional Neural Networks (CNNs) in computer vision. This empirical breakthrough is even more r…
cs.LG2021
Momentum Residual Neural Networks
Michael E. Sander, Pierre Ablin, Mathieu Blondel +1
The training of deep residual neural networks (ResNets) with backpropagation has a memory cost that increases linearly with respect to the depth of the network. A way to circumvent…