4 citations · 4 across the 3 of their papers we have counts for
1 paper · 1 filter
Han Tian, Luxuan Chen, Xinran Chen +10
Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and computational costs that make lo…