1 paper · 1 filter
Kai-Xuan Ding, Hao-Xiang Xu, Ji-Hua Peng +3
Steering with Sparse Autoencoders (SAEs) offers a lightweight inference-time path for adapting the behavior of large language models without retraining. By exposing sparse and inte…