1 paper
Elena Golimblevskaia, Aakriti Jain, Bruno Puri +3
The fields of explainable AI and mechanistic interpretability aim to uncover the internal structure of neural networks, with circuit discovery as a central tool for understanding m…