1 paper · 1 filter
Xiusi Chen, Hongzhi Wen, Sreyashi Nag +5
With the rapid development of large language models (LLMs), aligning LLMs with human values and societal norms to ensure their reliability and safety has become crucial. Reinforcem…