1 paper
Viacheslav Sinii, Alexey Gorbatovski, Artem Cherepanov +3
We show that training a single d-dimensional steering vector per layer with reinforcement learning, while freezing all base weights, matches the accuracy of fully RL-tuned reason…