1 paper
Prathyush Poduval, Calvin Yeung, Neel Desai +1
Sparse autoencoders are usually trained one layer at a time, even though transformer residual stream activations are strongly coupled across depth. This creates a practical problem…