From the 1 of 12 linked papers with an AI index.
3 papers · 1 filter
The Mosaic Memory of Large Language Models
Igor Shilov, Matthieu Meeus, Yves-Alexandre de Montjoye
As Large Language Models (LLMs) become widely adopted, understanding how they learn from, and memorize, training data becomes crucial. Memorization in LLMs is widely assumed to onl…
SoK: Membership Inference Attacks on LLMs are Rushing Nowhere (and How to Fix It)
Matthieu Meeus, Igor Shilov, Shubham Jain +3
Whether LLMs memorize their training data and what this means, from measuring privacy leakage to detecting copyright violations, has become a rapidly growing area of research. In t…
Copyright Traps for Large Language Models
Matthieu Meeus, Igor Shilov, Manuel Faysse +1
Questions of fair use of copyright-protected content to train Large Language Models (LLMs) are being actively debated. Document-level inference has been proposed as a new task: inf…