1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.LG2021
Thompson Sampling for Unimodal Bandits
Long Yang, Zhao Li, Zehong Hu +4
In this paper, we propose a Thompson Sampling algorithm for \emph{unimodal} bandits, where the expected reward is unimodal over the partially ordered arms. To exploit the unimodal…
cs.LG2019★ 1 cited
Which Channel to Ask My Question? Personalized Customer Service Request Stream Routing using Deep Reinforcement Learning
Zining Liu, Chong Long, Xiaolu Lu +3
Customer services are critical to all companies, as they may directly connect to the brand reputation. Due to a great number of customers, e-commerce companies often employ multipl…
cs.GT2018
Inference Aided Reinforcement Learning for Incentive Mechanism Design in Crowdsourcing
Zehong Hu, Yitao Liang, Yang Liu +1
Incentive mechanisms for crowdsourcing are designed to incentivize financially self-interested workers to generate and report high-quality labels. Existing mechanisms are often dev…