1 paper · 1 filter
Ole Jorgensen, Dylan Cope, Nandi Schoots +1
Recent work in activation steering has demonstrated the potential to better control the outputs of Large Language Models (LLMs), but it involves finding steering vectors. This is d…