2 citations · 2 across the 1 of their papers we have counts for
1 paper
Kanishk Gandhi, Denise Lee, Gabriel Grand +4
Language models are rarely shown fruitful mistakes while training. They then struggle to look beyond the next token, suffering from a snowballing of errors and struggling to predic…