13 citations · 17 across the 3 of their papers we have counts for
1 paper · 1 filter
David Owen
We investigate large language model performance across five orders of magnitude of compute scaling in eleven recent model architectures. We show that average benchmark performance,…