9 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.LG2020
Optimal Mixture Weights for Off-Policy Evaluation with Multiple Behavior Policies
Jinlin Lai, Lixin Zou, Jiaxing Song
Off-policy evaluation is a key component of reinforcement learning which evaluates a target policy with offline data collected from behavior policies. It is a crucial step towards…
cs.LG2019★ 9 cited
On the Necessity and Effectiveness of Learning the Prior of Variational Auto-Encoder
Haowen Xu, Wenxiao Chen, Jinlin Lai +3
Using powerful posterior distributions is a popular approach to achieving better variational inference. However, recent works showed that the aggregated posterior may fail to match…