1 paper
Jiyoung Lee, Seungho Kim, Seunghyun Won +6
AI alignment refers to models acting towards human-intended goals, preferences, or ethical principles. Given that most large-scale deep learning models act as black boxes and canno…