1 paper · 1 filter
Minjae Kang, Jaehyung Kim
Large Language Models (LLMs), despite advances in instruction tuning, often fail to follow complex user instructions. Activation steering techniques aim to mitigate this by manipul…