3 papers
cs.LG2026
Demystifying the Slash Pattern in Attention: The Role of RoPE
Yuan Cheng, Fengzhuo Zhang, Yunlong Hou +5
Large Language Models (LLMs) often exhibit slash attention patterns, where attention scores concentrate along the -th sub-diagonal for some offset . These patterns play a key…
cs.CL2025
Benchmarking and Rethinking Knowledge Editing for Large Language Models
Guoxiu He, Xin Song, Futing Wang +1
Knowledge editing aims to update the embedded knowledge within Large Language Models (LLMs). However, existing approaches, whether through parameter modification or external memory…
cs.CL2025
Knowledge Updating? No More Model Editing! Just Selective Contextual Reasoning
Guoxiu He, Xin Song, Aixin Sun
As real-world knowledge evolves, the information embedded within large language models (LLMs) can become outdated, inadequate, or erroneous. Model editing has emerged as a prominen…