1 paper
Siyan Zhao, Mingyi Hong, Yang Liu +2
Large Language Models (LLMs) are increasingly used as chatbots, yet their ability to personalize responses to user preferences remains limited. We introduce PrefEval, a benchmark f…