443 citations · 554 across the 16 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models
Sunny Rai, Jinyi Kuang, Reyhan Jamalova +7
Previous AI alignment efforts have focused primarily on first-order social norms -- teaching models what is socially acceptable or unacceptable (e.g., `do not steal'). However, soc…
cs.AI2024
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
Aadesh Salecha, Molly E. Ireland, Shashanka Subrahmanya +3
As Large Language Models (LLMs) become widely used to model and simulate human behavior, understanding their biases becomes critical. We developed an experimental framework using B…