1 paper · 1 filter
Parand A. Alamdari, Soroush Ebadian, Ariel D. Procaccia
We consider the challenge of AI value alignment with multiple individuals that have different reward functions and optimal policies in an underlying Markov decision process. We for…