1 paper · 1 filter
Raquib Bin Yousuf, Aadyant Khatri, Shengzhe Xu +2
Recently proposed evaluation benchmarks aim to characterize the effective context length and the forgetting tendencies of large language models (LLMs). However, these benchmarks of…