1 paper
Kristian González Barman, Simon Lohse, Henk de Regt
We argue for the epistemic and ethical advantages of pluralism in Reinforcement Learning from Human Feedback (RLHF) in the context of Large Language Models (LLM). Drawing on social…