1 paper · 1 filter
Ruihan Zhang, Jun Sun
Large language models (LLMs) are increasingly trained on massive, heterogeneous text corpora, raising serious concerns about the unauthorised use of proprietary or personal data du…