1 paper · 1 filter
Falko Helm, Nico Daheim, Iryna Gurevych
Many applications of large language models (LLMs) require long-context understanding, but models continue to struggle with such tasks. We hypothesize that conventional next-token p…