1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Addison J. Wu, Ryan Liu, Shuyue Stella Li +2
Large language models (LLMs) are trained to align with user preferences through methods like reinforcement learning. Yet models are beginning to be deployed not solely to satisfy u…