1 paper · 1 filter
Mingye Zhu, Yi Liu, Lei Zhang +2
Recently, tremendous strides have been made to align the generation of Large Language Models (LLMs) with human values to mitigate toxic or unhelpful content. Leveraging Reinforceme…