1 citations · 1 across the 2 of their papers we have counts for
2 papers
stat.ML2022★ 1 cited
Low-variance estimation in the Plackett-Luce model via quasi-Monte Carlo sampling
Alexander Buchholz, Jan Malte Lichtenberg, Giuseppe Di Benedetto +3
The Plackett-Luce (PL) model is ubiquitous in learning-to-rank (LTR) because it provides a useful and intuitive probabilistic model for sampling ranked lists. Counterfactual offlin…
cs.LG2019
Iterative Policy-Space Expansion in Reinforcement Learning
Jan Malte Lichtenberg, Özgür Şimşek
Humans and animals solve a difficult problem much more easily when they are presented with a sequence of problems that starts simple and slowly increases in difficulty. We explore…