2 papers
cs.CL2025
Robust Preference Optimization via Dynamic Target Margins
Jie Sun, Junkang Wu, Jiancan Wu +5
The alignment of Large Language Models (LLMs) is crucial for ensuring their safety and reliability in practical applications. Direct Preference Optimization (DPO) has emerged as an…
math.OC2025
OptScaler: A Collaborative Framework for Robust Autoscaling in the Cloud
Ding Zou, Wei Lu, Zhibo Zhu +7
Autoscaling is a critical mechanism in cloud computing, enabling the autonomous adjustment of computing resources in response to dynamic workloads. This is particularly valuable fo…