5 citations · 8 across the 3 of their papers we have counts for
2 papers
cs.AI2023★ 5 cited
Moral Foundations of Large Language Models
Marwa Abdulhai, Gregory Serapio-Garcia, Clément Crepy +3
Moral foundations theory (MFT) is a psychological assessment tool that decomposes human moral reasoning into five factors, including care/harm, liberty/oppression, and sanctity/deg…
cs.LG2022★ 2 cited
Basis for Intentions: Efficient Inverse Reinforcement Learning using Past Experience
Marwa Abdulhai, Natasha Jaques, Sergey Levine
This paper addresses the problem of inverse reinforcement learning (IRL) -- inferring the reward function of an agent from observing its behavior. IRL can provide a generalizable a…