1 citations · 1 across the 1 of their papers we have counts for
1 paper
Arash Ahmadian, Chris Cremer, Matthias Gallé +5
AI alignment in the shape of Reinforcement Learning from Human Feedback (RLHF) is increasingly treated as a crucial ingredient for high performance large language models. Proximal…