4 citations · 4 across the 2 of their papers we have counts for
1 paper · 1 filter
Yang Zhao, Yixin Wang, Mingzhang Yin
Aligning Large Language Models (LLMs) with human preferences is crucial in ensuring desirable and controllable model behaviors. Current methods, such as Reinforcement Learning from…