2 papers
cs.LG2025
Pr{é}diction optimale pour un mod{è}le ordinal {à} covariables fonctionnelles
Simón Weinberger, Jairo Cugliari, Aurélie Le Cain
We present a prediction framework for ordinal models: we introduce optimal predictions using loss functions and give the explicit form of the Least-Absolute-Deviation prediction fo…
cs.LG2025
Policy gradient methods for ordinal policies
Simón Weinberger, Jairo Cugliari
In reinforcement learning, the softmax parametrization is the standard approach for policies over discrete action spaces. However, it fails to capture the order relationship betwee…