5 citations · 5 across the 1 of their papers we have counts for
1 paper · 1 filter
Dongyoung Go, Tomasz Korbak, Germán Kruszewski +3
Aligning language models with preferences can be posed as approximating a target distribution representing some desired behavior. Existing approaches differ both in the functional…