1 paper · 1 filter
Zhenyu Ding, Yuhao Wang, Tengyue Xiao +3
Large Language Models (LLMs) demonstrate impressive capabilities, yet their outputs often suffer from misalignment with human preferences due to the inadequacy of weak supervision…