1 paper
Zoë Prins, Samuele Punzo, Frank Wildenburg +2
Standard evaluations of Large language models (LLMs) focus on task performance, offering limited insight into whether correct behavior reflects appropriate underlying mechanisms an…