1 paper
Gilles Bareilles, Julien Fageot, Lê-Nguyên Hoang +4
Comparison-based preference learning has become central to the alignment of AI models with human preferences. However, these methods may behave counterintuitively. After empiricall…