2 papers
cs.LG2021
Thompson Sampling for Gaussian Entropic Risk Bandits
Ming Liang Ang, Eloise Y. Y. Lim, Joel Q. L. Chang
The multi-armed bandit (MAB) problem is a ubiquitous decision-making problem that exemplifies exploration-exploitation tradeoff. Standard formulations exclude risk in decision maki…
cs.LG2020
Refining Deep Generative Models via Discriminator Gradient Flow
Abdul Fatir Ansari, Ming Liang Ang, Harold Soh
Deep generative modeling has seen impressive advances in recent years, to the point where it is now commonplace to see simulated samples (e.g., images) that closely resemble real-w…