1 paper
Sunny Rai, Jinyi Kuang, Reyhan Jamalova +7
Previous AI alignment efforts have focused primarily on first-order social norms -- teaching models what is socially acceptable or unacceptable (e.g., `do not steal'). However, soc…