1 paper
Ling Zhang, Xianliang Yang, Juwon Yu +4
Fine-tuning large pretrained language models is a common approach for aligning them with human preferences, but noisy or off-target examples can dilute supervision. While small, we…