2 citations · 2 across the 3 of their papers we have counts for
3 papers · 1 filter
Deep Exploration with PAC-Bayes
Bahareh Tasdighi, Manuel Haussmann, Nicklas Werge +2
Reinforcement learning (RL) for continuous control under delayed rewards is an under-explored problem despite its significance in real-world applications. Many complex skills are b…
Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits
Nicklas Werge, Yi-Shan Wu, Abdullah Akgül +1
In non-stationary linear contextual bandits, existing efficient algorithms typically rely on the Weighted Regularized Least-Squares (WRLS) estimator. Because WRLS only provides poi…
PAC-Bayesian Soft Actor-Critic Learning
Bahareh Tasdighi, Abdullah Akgül, Manuel Haussmann +2
Actor-critic algorithms address the dual goals of reinforcement learning (RL), policy evaluation and improvement via two separate function approximators. The practicality of this a…