2 papers
cs.LG2026
Knowledge Offloading: Decomposing LLMs into Sparse Backbones and Memory Modules
Karim Galliamov, Rochelle Choenni, Ivan Titov
LLMs encode both general capabilities and domain-specific knowledge in a single set of parameters. We ask whether this capacity can be reorganized: keeping broadly useful computati…
cs.CL2024
Generalisation First, Memorisation Second? Memorisation Localisation for Natural Language Classification Tasks
Verna Dankers, Ivan Titov
Memorisation is a natural part of learning from real-world data: neural models pick up on atypical input-output combinations and store those training examples in their parameter sp…