Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
RS-DPO: A Hybrid Rejection Sampling and Direct Preference Optimization Method for Alignment of Large Language Models
Saeed Khaki, JinJin Li, Lan Ma +2
Reinforcement learning from human feedback (RLHF) has been extensively employed to align large language models with user intent. However, proximal policy optimization (PPO) based R…
cs.CL2023
Uncovering Drift in Textual Data: An Unsupervised Method for Detecting and Mitigating Drift in Machine Learning Models
Saeed Khaki, Akhouri Abhinav Aditya, Zohar Karnin +3
Drift in machine learning refers to the phenomenon where the statistical properties of data or context, in which the model operates, change over time leading to a decrease in its p…