5 citations · 5 across the 5 of their papers we have counts for
1 paper · 1 filter
Wassim Bouaziz, Mathurin Videau, Nicolas Usunier +1
The pre-training of large language models (LLMs) relies on massive text datasets sourced from diverse and difficult-to-curate origins. Although membership inference attacks and hid…