1 paper
Rebonto Haque, Oliver M. Turnbull, Anisha Parsan +4
Sparse autoencoders (SAEs) are a mechanistic interpretability technique that have been used to provide insight into learned concepts within large protein language models. Here, we…