1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Louis Castricato, Nathan Lile, Suraj Anand +3
Existing methods for controlling language models, such as RLHF and Constitutional AI, involve determining which LLM behaviors are desirable and training them into a language model.…