1 paper
Rafael Rafailov, Archit Sharma, Eric Mitchell +3
While large-scale unsupervised language models (LMs) learn broad world knowledge and some reasoning skills, achieving precise control of their behavior is difficult due to the comp…