1 paper
Xu Chu, Zhixin Zhang, Tianyu Jia +1
Aligning large language models (LLMs) with human preferences typically demands vast amounts of meticulously curated data, which is both expensive and prone to labeling noise. We pr…