1 paper
Quan Xiao, Yutong Xuan, Gaowen Liu +2
Supervised fine-tuning (SFT) datasets are critical to the downstream performance of large language models, yet they often contain low-quality or harmful question-response pairs. To…