1 paper
Walter Nelson, Theofanis Karaletsos, Francesco Locatello
Recently, sparse autoencoders (SAEs) have emerged as an attractive tool for interpreting and interacting with representations in practical neural networks. While it is common empir…