3 papers
cs.AI2026
CORDA: A Benchmark for Hierarchical Harm-Centric Moral Reasoning in Large Language Models
Siddarth Singh, Victoria Williams, Simon Rosen +6
The key question in moral judgement is not simply whether someone chooses the "right" answer, but how they decide what matters most when moral principles conflict. Current evaluati…
cs.AI2026
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
Simon Rosen, Siddarth Singh, Ebenezer Gelo +6
Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection of AI safety, moral philosophy, and c…
cs.CL2025
Heartificial Intelligence: Exploring Empathy in Language Models
Victoria Williams, Benjamin Rosman
Large language models have become increasingly common, used by millions of people worldwide in both professional and personal contexts. As these models continue to advance, they ar…