1 paper · 1 filter
Zetian Sun, Dongfang Li, Xuhui Chen +2
The alignment of language models~(LMs) with human preferences is critical for building reliable AI systems. The problem is typically framed as optimizing an LM policy to maximize t…