1 paper
Ashwin Saraswatula, David Klindt
Sparse autoencoders (SAEs) have emerged as a promising approach for learning interpretable features from neural network activations. However, the optimization landscape for SAE tra…