1 paper · 1 filter
Zhuojun Gu, Quan Wang, Shuchu Han
Recent advances in Large Language Models (LLMs) highlight the need to align their behaviors with human values. A critical, yet understudied, issue is the potential divergence betwe…