3 citations · 3 across the 1 of their papers we have counts for
1 paper · 1 filter
David Owen
We investigate large language model performance across five orders of magnitude of compute scaling in eleven recent model architectures. We show that average benchmark performance,…