1 citations · 1 across the 1 of their papers we have counts for
1 paper
Louis Castricato, Nathan Lile, Suraj Anand +3
Existing methods for controlling language models, such as RLHF and Constitutional AI, involve determining which LLM behaviors are desirable and training them into a language model.…