1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Katie Z Luo, Zhenzhen Liu, Xiangyu Chen +7
Recent advances in machine learning have shown that Reinforcement Learning from Human Feedback (RLHF) can improve machine learning models and align them with human preferences. Alt…