1 citations · 1 across the 2 of their papers we have counts for
3 papers
Gotta Catch them all: the modes of Sycophancy
Shreyans Jain, Alexandra Yost, Amirali Abdullah
Large language models often align with users' beliefs at the expense of factual accuracy, a behavior known as sycophancy. Prior mechanistic studies largely treat sycophancy as a si…
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
Alexandra Yost, Shreyans Jain, Shivam Raval +6
Persona conditioning is widely used to steer large language model (LLM) behavior, but it is unclear whether it induces stable behavioral structure or superficial variation. We prop…
Sycophancy as compositions of Atomic Psychometric Traits
Shreyans Jain, Alexandra Yost, Amirali Abdullah
Sycophancy is a key behavioral risk in LLMs, yet is often treated as an isolated failure mode that occurs via a single causal mechanism. We instead propose modeling it as geometric…