most citedHRLAIF: Improvements in Helpfulness and Harmlessness in Open-domain Reinforcement Learning From AI Feedback

3 citations · 3 across the 4 of their papers we have counts for

collaborators

4 papers