1 citations · 1 across the 1 of their papers we have counts for
1 paper
Katie Z Luo, Zhenzhen Liu, Xiangyu Chen +7
Recent advances in machine learning have shown that Reinforcement Learning from Human Feedback (RLHF) can improve machine learning models and align them with human preferences. Alt…