1 citations · 1 across the 1 of their papers we have counts for
1 paper
Marcus Williams
This paper presents Multi-Objective Reinforcement Learning from AI Feedback (MORLAIF), a novel approach to improving the alignment and performance of language models trained using…