Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Morality is Contextual: Learning Interpretable Moral Contexts from Human Data with Probabilistic Clustering and Large Language Models
Geoffroy Morlat, Marceau Nahon, Augustin Chartouny +3
Moral actions are judged not only by their outcomes but by the context in which they occur. We present COMETH (Contextual Organization of Moral Evaluation from Textual Human inputs…
cs.CL2025
Semantic Deception: When Reasoning Models Can't Compute an Addition
Nathaniël de Leeuw, Marceau Nahon, Mathis Reymond +2
Large language models (LLMs) are increasingly used in situations where human values are at stake, such as decision-making tasks that involve reasoning when performed by humans. We…
cs.CL2024
Strong and weak alignment of large language models with human values
Mehdi Khamassi, Marceau Nahon, Raja Chatila
Minimizing negative impacts of Artificial Intelligent (AI) systems on human societies without human supervision requires them to be able to align with human values. However, most c…