1 paper · 1 filter
Charlie Snell, Eric Wallace, Dan Klein +1
A fundamental open challenge in modern LLM scaling is the lack of understanding around emergent capabilities. In particular, language model pretraining loss is known to be highly p…