1 paper · 1 filter
Aidan Kierans, Ritam Dutt, Kaley Rittichier +2
Recent work on evaluating the moral competence of large language models (LLMs) has focused primarily on what we call the moral value problem, i.e., whether model outputs align with…