1 citations · 1 across the 1 of their papers we have counts for
1 paper
Addison J. Wu, Ryan Liu, Shuyue Stella Li +2
Large language models (LLMs) are trained to align with user preferences through methods like reinforcement learning. Yet models are beginning to be deployed not solely to satisfy u…