1 paper
Jérémy Scheurer, Jon Ander Campos, Tomasz Korbak +4
Pretrained language models often generate outputs that are not in line with human preferences, such as harmful text or factually incorrect summaries. Recent work approaches the abo…