42 citations · 94 across the 7 of their papers we have counts for
3 papers · 1 filter
A Scalable MIP-based Method for Learning Optimal Multivariate Decision Trees
Haoran Zhu, Pavankumar Murali, Dzung T. Phan +2
Several recent publications report advances in training optimal decision trees (ODT) using mixed-integer programs (MIP), due to algorithmic advances in integer programming and a gr…
A Hybrid Stochastic Policy Gradient Algorithm for Reinforcement Learning
Nhan H. Pham, Lam M. Nguyen, Dzung T. Phan +3
We propose a novel hybrid stochastic policy gradient estimator by combining an unbiased policy gradient estimator, the REINFORCE estimator, with another biased one, an adapted SARA…
DTN: A Learning Rate Scheme with Convergence Rate of for SGD
Lam M. Nguyen, Phuong Ha Nguyen, Dzung T. Phan +2
This paper has some inconsistent results, i.e., we made some failed claims because we did some mistakes for using the test criterion for a series. Precisely, our claims on the conv…