2 citations · 4 across the 12 of their papers we have counts for
1 paper · 1 filter
Roy Rinberg, Usha Bhalla, Igor Shilov +2
Targeted interventions on language models, such as unlearning or model editing, aim to modify specific information, but their effects often propagate to related, unintended areas (…