5 citations · 8 across the 6 of their papers we have counts for
1 paper · 1 filter
Kyuyoung Kim, Ah Jeong Seo, Hao Liu +2
Large language models (LLMs) fine-tuned with alignment techniques, such as reinforcement learning from human feedback, have been instrumental in developing some of the most capable…