12 citations · 23 across the 4 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020★ 8 cited
Hierarchical Adaptive Contextual Bandits for Resource Constraint based Recommendation
Mengyue Yang, Qingyang Li, Zhiwei Qin +1
Contextual multi-armed bandit (MAB) achieves cutting-edge performance on a variety of problems. When it comes to real-world scenarios such as recommendation system and online adver…
cs.LG2019★ 3 cited
Environment Reconstruction with Hidden Confounders for Reinforcement Learning based Recommendation
Wenjie Shang, Yang Yu, Qingyang Li +3
Reinforcement learning aims at searching the best policy model for decision making, and has been shown powerful for sequential recommendations. The training of the policy by reinfo…