483 citations · 555 across the 13 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Investigating Model Editing for Unlearning in Large Language Models
Shariqah Hossain, Lalana Kagal
Machine unlearning aims to remove unwanted information from a model, but many methods are inefficient for LLMs with large numbers of parameters or fail to fully remove the intended…
cs.CL2024★ 1 cited
Towards Resource Efficient and Interpretable Bias Mitigation in Large Language Models
Schrasing Tong, Eliott Zemour, Jessica Lu +2
Although large language models (LLMs) have demonstrated their effectiveness in a wide range of applications, they have also been observed to perpetuate unwanted biases present in t…