1 paper · 1 filter
Ziyun Cui, Ziyang Zhang, Guangzhi Sun +2
Advances in large language models raise the question of how alignment techniques will adapt as models become increasingly complex and humans will only be able to supervise them wea…