1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Proximal Ranking Policy Optimization for Practical Safety in Counterfactual Learning to Rank
Shashank Gupta, Harrie Oosterhuis, Maarten de Rijke
Counterfactual learning to rank (CLTR) can be risky and, in various circumstances, can produce sub-optimal models that hurt performance when deployed. Safe CLTR was introduced to m…
cs.IR2024★ 1 cited
A First Look at Selection Bias in Preference Elicitation for Recommendation
Shashank Gupta, Harrie Oosterhuis, Maarten de Rijke
Preference elicitation explicitly asks users what kind of recommendations they would like to receive. It is a popular technique for conversational recommender systems to deal with…