6 citations · 6 across the 2 of their papers we have counts for
1 paper · 1 filter
Jatin Nainani, Sankaran Vaidyanathan, AJ Yeung +2
Mechanistic interpretability aims to understand the inner workings of large neural networks by identifying circuits, or minimal subgraphs within the model that implement algorithms…