3 papers
cs.LG2022
Policy Learning for Robust Markov Decision Process with a Mismatched Generative Model
Jialian Li, Tongzheng Ren, Dong Yan +2
In high-stake scenarios like medical treatment and auto-piloting, it's risky or even infeasible to collect online experimental data to train the agent. Simulation-based training ca…
cs.LG2018
Lazy-CFR: fast and near optimal regret minimization for extensive games with imperfect information
Yichi Zhou, Tongzheng Ren, Jialian Li +2
Counterfactual regret minimization (CFR) is the most popular algorithm on solving two-player zero-sum extensive games with imperfect information and achieves state-of-the-art perfo…
cs.LG2016
Conditional Generative Moment-Matching Networks
Yong Ren, Jialian Li, Yucen Luo +1
Maximum mean discrepancy (MMD) has been successfully applied to learn deep generative models for characterizing a joint distribution of variables via kernel mean embedding. In this…