1 paper · 1 filter
Samarth Raina, Saksham Aggarwal, Aman Chadha +2
Direct Preference Optimization (DPO) has become a standard recipe for aligning large language models, yet it is still unclear what kind of change it actually induces inside the net…