1 paper
Oscar Gilg, Pierre Beckmann, Daniel Paleka +1
Large language models (LLMs) can be said to have preferences: they reliably pick certain tasks and outputs over others, and preferences shaped by post-training and system prompts a…