1 paper
Yoshihiro Izawa, Gouki Minegishi, Koshi Eguchi +2
Activation steering offers a computationally efficient mechanism for controlling Large Language Models (LLMs) without fine-tuning. While effectively controlling target traits (e.g.…