1 paper · 1 filter
Seongho Son, William Bankes, Sayak Ray Chowdhury +2
Current Large Language Model (LLM) preference optimization algorithms do not account for temporal preference drift, which can lead to severe misalignment. To address this limitatio…