1 paper · 1 filter
Junlong Li, Fan Zhou, Shichao Sun +3
As a relative quality comparison of model responses, human and Large Language Model (LLM) preferences serve as common alignment goals in model fine-tuning and criteria in evaluatio…