Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning
Siddharth Aravindan, Dixant Mittal, Wee Sun Lee
Posterior Sampling for Reinforcement Learning (PSRL) is a well-known algorithm that augments model-based reinforcement learning (MBRL) algorithms with Thompson sampling. PSRL maint…
cs.LG2024
Continual Learning of Numerous Tasks from Long-tail Distributions
Liwei Kang, Wee Sun Lee
Continual learning, an important aspect of artificial intelligence and machine learning research, focuses on developing models that learn and adapt to new tasks while retaining pre…
cs.LG2024
Differentiable Tree Search Network
Dixant Mittal, Wee Sun Lee
In decision-making problems with limited training data, policy functions approximated using deep neural networks often exhibit suboptimal performance. An alternative approach invol…