2 papers
cs.LG2026
DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable)
Wenxuan Zhou, Shujian Zhang, Brice Magdalou +4
Normative theories allow one to elicit key parts of a ML algorithm from first principles, which is crucial at a time of championed scrutiny for ML work. Direct Preference Optimizat…
cs.LG2024
Generative Forests
Richard Nock, Mathieu Guillame-Bert
We focus on generative AI for a type of data that still represent one of the most prevalent form of data: tabular data. Our paper introduces two key contributions: a new powerful c…