33 citations · 36 across the 5 of their papers we have counts for
Showing 2026Show all
3 papers · 1 filter
cs.AI2026
Rating the Raters: Rasch Measurement Theory for LLM Evaluation
Pratik S. Sachdeva, Nathan Boudol
LLMs now sit on every side of evaluation: as examinees scored on benchmarks, judges of other models' outputs, and raters of human-generated content. Each paradigm can be viewed as…
cs.CY2026
The Fabricated Front: Generative AI and the Opacity of Workplace Performance
Tom van Nuenen, Pratik S. Sachdeva, Sahiba Chopra
Generative AI (GenAI) has become a fixture of workplace life. Current research asks chiefly what this implies for jobs and outputs, measured in productivity, displacement, or bias.…
cs.CL2026
The Fragility Of Moral Judgment In Large Language Models
Tom van Nuenen, Pratik S. Sachdeva
People increasingly use large language models (LLMs) for everyday moral and interpersonal guidance, yet these systems cannot interrogate missing context and judge dilemmas as prese…