108 citations · 128 across the 4 of their papers we have counts for
1 paper · 1 filter
Alexander Maloney, Daniel A. Roberts, James Sully
Large language models with a huge number of parameters, when trained on near internet-sized number of tokens, have been empirically shown to obey neural scaling laws: specifically,…