2 citations · 2 across the 1 of their papers we have counts for
1 paper
Wes Gurnee, Theo Horsley, Zifan Carl Guo +5
A basic question within the emerging field of mechanistic interpretability is the degree to which neural networks learn the same underlying mechanisms. In other words, are neural m…