1 citations · 1 across the 1 of their papers we have counts for
1 paper
Xiaobei Wang, Shuchang Liu, Qingpeng Cai +4
Recent advances in recommender systems have shown that user-system interaction essentially formulates long-term optimization problems, and online reinforcement learning can be adop…