1 paper
Martin Spišák, Ladislav Peška, Petr Škoda +2
Sparse autoencoders (SAEs) have recently emerged as pivotal tools for introspection into large language models. SAEs can uncover high-quality, interpretable features at different l…