6 citations · 6 across the 3 of their papers we have counts for
1 paper · 1 filter
Miao Fan, Chen Hu, Shuchang Zhou
The Reinforcement Learning from Human Feedback (RLHF) plays a pivotal role in shaping the impact of large language models (LLMs), contributing significantly to controlling output t…