1 paper
Ruixuan Deng, Xiaoyang Hu, Miles Gilberti +5
We identify semantically coherent, context-consistent network components in large language models (LLMs) using coactivation of sparse autoencoder (SAE) features collected from just…