3 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 3 cited
Training a Vision Transformer from scratch in less than 24 hours with 1 GPU
Saghar Irandoust, Thibaut Durand, Yunduz Rakhmangulova +2
Transformers have become central to recent advances in computer vision. However, training a vision Transformer (ViT) model from scratch can be resource intensive and time consuming…
cs.CL2021
Turing: an Accurate and Interpretable Multi-Hypothesis Cross-Domain Natural Language Database Interface
Peng Xu, Wenjie Zi, Hamidreza Shahidi +7
A natural language database interface (NLDB) can democratize data-driven insights for non-technical users. However, existing Text-to-SQL semantic parsers cannot achieve high enough…
cs.CL2020
Optimizing Deeper Transformers on Small Datasets
Peng Xu, Dhruv Kumar, Wei Yang +6
It is a common belief that training deep transformers from scratch requires large datasets. Consequently, for small datasets, people usually use shallow and simple additional layer…