1 paper
Youxiang Zhu, Ruochen Li, Danqing Wang +2
Long-context large language models (LLMs) are prone to be distracted by irrelevant contexts. The reason for distraction remains poorly understood. In this paper, we first identify…