9.8k citations · 18.7k across the 31 of their papers we have counts for
1 paper · 1 filter
Yufeng Du, Minyang Tian, Srikanth Ronanki +7
Large language models (LLMs) often fail to scale their performance on long-context tasks performance in line with the context lengths they support. This gap is commonly attributed…