42 citations · 94 across the 15 of their papers we have counts for
Showing 2020Show all
2 papers · 1 filter
cs.LG2020★ 17 cited
A Scalable MIP-based Method for Learning Optimal Multivariate Decision Trees
Haoran Zhu, Pavankumar Murali, Dzung T. Phan +2
Several recent publications report advances in training optimal decision trees (ODT) using mixed-integer programs (MIP), due to algorithmic advances in integer programming and a gr…
cs.LG2020
A Hybrid Stochastic Policy Gradient Algorithm for Reinforcement Learning
Nhan H. Pham, Lam M. Nguyen, Dzung T. Phan +3
We propose a novel hybrid stochastic policy gradient estimator by combining an unbiased policy gradient estimator, the REINFORCE estimator, with another biased one, an adapted SARA…