1 paper · 1 filter
Victor Lecomte, Kushal Thaman, Rylan Schaeffer +3
Polysemantic neurons -- neurons that activate for a set of unrelated features -- have been seen as a significant obstacle towards interpretability of task-optimized deep networks,…