1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.LG2022★ 1 cited
Scalable Training of Language Models using JAX pjit and TPUv4
Joanna Yoo, Kuba Perlin, Siddhartha Rao Kamalakara +1
Modern large language models require distributed training strategies due to their size. The challenges of efficiently and robustly training them are met with rapid developments on…
cs.LG2020
Improving compute efficacy frontiers with SliceOut
Pascal Notin, Aidan N. Gomez, Joanna Yoo +1
Pushing forward the compute efficacy frontier in deep learning is critical for tasks that require frequent model re-training or workloads that entail training a large number of mod…