1 paper · 1 filter
Hung Le, Quan Tran, Dung Nguyen +4
How can Large Language Models (LLMs) be aligned with human intentions and values? A typical solution is to gather human preference on model outputs and finetune the LLMs accordingl…