1 paper · 1 filter
Will LeVine, Benjamin Pikus, Anthony Chen +1
Foundation models, specifically Large Language Models (LLMs), have lately gained wide-spread attention and adoption. Reinforcement Learning with Human Feedback (RLHF) involves trai…