1 paper
Hussein Chouman, Wataru Sasaki, Tomokazu Matsui +2
Mechanistic interpretability has produced a rich inventory of component-level analyses that characterise what neural-network components encode and how they interact. Their outputs,…