1 paper · 1 filter
Ashwin Saraswatula, David Klindt
Sparse autoencoders (SAEs) have emerged as a promising approach for learning interpretable features from neural network activations. However, the optimization landscape for SAE tra…