1 paper
Junlong Li, Fan Zhou, Shichao Sun +3
As a relative quality comparison of model responses, human and Large Language Model (LLM) preferences serve as common alignment goals in model fine-tuning and criteria in evaluatio…