2 papers
cs.CL2025
Triple Preference Optimization: Achieving Better Alignment using a Single Step Optimization
Amir Saeidi, Shivanshu Verma, Aswin RRV +2
Reinforcement Learning with Human Feedback (RLHF) enhances the alignment of Large Language Models (LLMs). However, its limitations have led to the development of Direct Preference…
cs.CL2025
Insights into Alignment: Evaluating DPO and its Variants Across Multiple Tasks
Amir Saeidi, Shivanshu Verma, Md Nayem Uddin +1
This study evaluates Direct Preference Optimization (DPO) and its variants for aligning Large Language Models (LLMs) with human preferences, testing three configurations: (1) with…