465 citations · 743 across the 21 of their papers we have counts for
1 paper · 1 filter
Kyuyoung Kim, Ah Jeong Seo, Hao Liu +2
Large language models (LLMs) fine-tuned with alignment techniques, such as reinforcement learning from human feedback, have been instrumental in developing some of the most capable…