Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution
Muhammad Umair Haider, Hammad Rizwan, Hassan Sajjad +2
Pervasive polysemanticity in large language models (LLMs) undermines discrete neuron-concept attribution, posing a significant challenge for model interpretation and control. We sy…
cs.LG2025
Instance-Level Difficulty: A Missing Perspective in Machine Unlearning
Hammad Rizwan, Mahtab Sarvmaili, Hassan Sajjad +1
Current research on deep machine unlearning primarily focuses on improving or evaluating the overall effectiveness of unlearning methods while overlooking the varying difficulty of…