From the 1 of 1 linked paper with an AI index.
1 paper
Binwen Liu, Yilin Ren
The paper introduces ESFP, a benchmark that tests whether large language models can shift between neutral attribution and personal stance when prompted differently, and evaluates t…