From the 1 of 8 linked papers with an AI index.
1 paper · 1 filter
Pankayaraj Pathmanathan, Furong Huang
Reward modeling (RM), which captures human preferences to align large language models (LLMs), is increasingly employed in tasks such as model finetuning, response filtering, and ra…