1 paper
Binwen Liu, Yilin Ren
The paper introduces ESFP, a benchmark that tests whether large language models can shift between neutral attribution and personal stance when prompted differently, and evaluates t…