3 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 2 cited
Patch Gradient Descent: Training Neural Networks on Very Large Images
Deepak K. Gupta, Gowreesh Mago, Arnav Chavan +1
Traditional CNN models are trained and tested on relatively low resolution images (<300 px), and cannot be directly operated on large-scale images due to compute and memory constra…
cs.CV2022★ 3 cited
Vision Transformer Slimming: Multi-Dimension Searching in Continuous Optimization Space
Arnav Chavan, Zhiqiang Shen, Zhuang Liu +3
This paper explores the feasibility of finding an optimal sub-model from a vision transformer and introduces a pure vision transformer slimming (ViT-Slim) framework. It can search…