5 citations · 5 across the 2 of their papers we have counts for
1 paper · 2 filters
Xinyu Li, Ruiyang Zhou, Zachary C. Lipton +1
Personalized large language models (LLMs) are designed to tailor responses to individual user preferences. While Reinforcement Learning from Human Feedback (RLHF) is a commonly use…