21 citations · 21 across the 3 of their papers we have counts for
4 papers
Optimal Mixture Weights for Off-Policy Evaluation with Multiple Behavior Policies
Jinlin Lai, Lixin Zou, Jiaxing Song
Off-policy evaluation is a key component of reinforcement learning which evaluates a target policy with offline data collected from behavior policies. It is a crucial step towards…
MIN: Co-Governing Multi-Identifier Network Architecture and its Prototype on Operator's Network
Hui Li, Jiangxing Wu, Xin Yang +28
IP protocol is the core of TCP/IP network layer. However, since IP address and its Domain Name are allocated and managed by a single agency, there are risks of centralization. The…
The Prototype of Decentralized Multilateral Co-Governing Post-IP Internet Architecture and Its Testing on Operator Networks
Hui Li, Jiangxing Wu, Kaixuan Xing +32
The Internet has become the most important infrastructure of modern society, while the existing IP network is unable to provide high-quality service. The unilateralism IP network i…
Reinforcement Learning to Optimize Long-term User Engagement in Recommender Systems
Lixin Zou, Long Xia, Zhuoye Ding +3
Recommender systems play a crucial role in our daily lives. Feed streaming mechanism has been widely used in the recommender system, especially on the mobile Apps. The feed streami…