1 paper · 1 filter
Chenxin An, Fei Huang, Jun Zhang +4
The ability of Large Language Models (LLMs) to process and generate coherent text is markedly weakened when the number of input tokens exceeds their pretraining length. Given the e…