4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CL2024
Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods
Jai Doshi, Asa Cooper Stickland
Large language model unlearning aims to remove harmful information that LLMs have learnt to prevent their use for malicious purposes. LLMU and RMU have been proposed as two methods…
cs.CY2024★ 4 cited
Sleeper Social Bots: a new generation of AI disinformation bots are already a political threat
Jaiv Doshi, Ines Novacic, Curtis Fletcher +7
This paper presents a study on the growing threat of "sleeper social bots," AI-driven social bots in the political landscape, created to spread disinformation and manipulate public…