492 citations · 492 across the 2 of their papers we have counts for
1 paper · 1 filter
Bart Bussmann, Noa Nabeshima, Adam Karvonen +1
Sparse autoencoders (SAEs) have emerged as a powerful tool for interpreting neural networks by extracting the concepts represented in their activations. However, choosing the size…