4 citations · 8 across the 12 of their papers we have counts for
1 paper · 2 filters
Yang Zhao, Yixin Wang, Mingzhang Yin
Aligning Large Language Models (LLMs) with human preferences is crucial in ensuring desirable and controllable model behaviors. Current methods, such as Reinforcement Learning from…