18 citations · 21 across the 8 of their papers we have counts for
2 papers
cs.LG2021★ 1 cited
Balanced Q-learning: Combining the Influence of Optimistic and Pessimistic Targets
Thommen George Karimpanal, Hung Le, Majid Abdolshah +4
The optimistic nature of the Q-learning target leads to an overestimation bias, which is an inherent problem associated with standard learning. Such a bias fails to account for…
cs.LG2020
Randomised Gaussian Process Upper Confidence Bound for Bayesian Optimisation
Julian Berk, Sunil Gupta, Santu Rana +1
In order to improve the performance of Bayesian optimisation, we develop a modified Gaussian process upper confidence bound (GP-UCB) acquisition function. This is done by sampling…