2 papers
cs.IR2019
Variance Reduction in Gradient Exploration for Online Learning to Rank
Huazheng Wang, Sonwoo Kim, Eric McCord-Snook +2
Online Learning to Rank (OL2R) algorithms learn from implicit user feedback on the fly. The key of such algorithms is an unbiased estimation of gradients, which is often (trivially…
cs.IR2018
Efficient Exploration of Gradient Space for Online Learning to Rank
Huazheng Wang, Ramsey Langley, Sonwoo Kim +2
Online learning to rank (OL2R) optimizes the utility of returned search results based on implicit feedback gathered directly from users. To improve the estimates, OL2R algorithms e…