From the 1 of 1 linked paper with an AI index.
1 paper
Yuhang Zhu, Mingxuan Du, Benfeng Xu +3
The paper presents PALATE, a benchmark that uses per‑user simulated agents to evaluate role‑playing language models through free‑form, multi‑turn conversations and personalized sat…