136 citations · 136 across the 1 of their papers we have counts for
1 paper
Sam McCandlish, Jared Kaplan, Dario Amodei +1
In an increasing number of domains it has been demonstrated that deep learning models can be trained using relatively large batch sizes without sacrificing data efficiency. However…