2 papers
cs.LG2025
EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning
Siddharth Aravindan, Dixant Mittal, Wee Sun Lee
Posterior Sampling for Reinforcement Learning (PSRL) is a well-known algorithm that augments model-based reinforcement learning (MBRL) algorithms with Thompson sampling. PSRL maint…
cs.LG2024
Differentiable Tree Search Network
Dixant Mittal, Wee Sun Lee
In decision-making problems with limited training data, policy functions approximated using deep neural networks often exhibit suboptimal performance. An alternative approach invol…