1 paper
Alexander Zap, Tobias Joppen, Johannes Fürnkranz
Reinforcement learning usually makes use of numerical rewards, which have nice properties but also come with drawbacks and difficulties. Using rewards on an ordinal scale (ordinal…