1 citations · 1 across the 3 of their papers we have counts for
1 paper · 2 filters
Sarah Pan
Large decoder-based language models have become the dominant architecture for reward modeling in reinforcement learning from human feedback (RLHF). However, as reward models are in…