2 papers
cs.LG2026
Direct Preference Optimization with Rating Information: Practical Algorithms and Provable Gains
Luca Viano, Ruida Zhou, Yifan Sun +4
The class of direct preference optimization (DPO) algorithms has emerged as a promising approach for solving the alignment problem in foundation models. These algorithms work with…
q-fin.PM2025
Representation of forward performance criteria with random endowment via FBSDE and its application to forward optimized certainty equivalent
Gechun Liang, Yifan Sun, Thaleia Zariphopoulou
We extend the notion of forward performance criteria to settings with random endowment in incomplete markets. Building on these results, we introduce and develop the novel concept…