1 paper · 1 filter
Nils Feldhus, Laura Kopf
Understanding the decision-making processes of neural networks is a central goal of mechanistic interpretability. In the context of Large Language Models (LLMs), this involves unco…