2 citations · 3 across the 3 of their papers we have counts for
1 paper · 2 filters
Asael Sorensen, Charles Brock, David Chamberlain +3
Mechanistic interpretability seeks to make verifiable statements about the internal behavior of large language models (LLMs). Many interpretability techniques struggle to scale wit…