33 citations · 36 across the 4 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Rating the Raters: Rasch Measurement Theory for LLM Evaluation
Pratik S. Sachdeva, Nathan Boudol
LLMs now sit on every side of evaluation: as examinees scored on benchmarks, judges of other models' outputs, and raters of human-generated content. Each paradigm can be viewed as…
cs.AI2025
Interaction Protocol Shapes Moral Judgment in Multi-Agent Debate
Pratik S. Sachdeva, Tom van Nuenen
As agentic AI systems are deployed in advisory and evaluative roles, understanding how multi-agent interactions shape behavior becomes essential. Multi-agent debate has been studie…
cs.AI2025★ 3 cited
Normative Evaluation of Large Language Models with Everyday Moral Dilemmas
Pratik S. Sachdeva, Tom van Nuenen
The rapid adoption of large language models (LLMs) has spurred extensive research into their encoded moral norms and decision-making processes. Much of this research relies on prom…