8 citations · 13 across the 2 of their papers we have counts for
3 papers
cs.LG2020
Expert Selection in High-Dimensional Markov Decision Processes
Vicenc Rubies-Royo, Eric Mazumdar, Roy Dong +2
In this work we present a multi-armed bandit framework for online expert selection in Markov decision processes and demonstrate its use in high-dimensional settings. Our method tak…
eess.SY2017★ 8 cited
A Multi-Armed Bandit Approach for Online Expert Selection in Markov Decision Processes
Eric Mazumdar, Roy Dong, Vicenç Rúbies Royo +2
We formulate a multi-armed bandit (MAB) approach to choosing expert policies online in Markov decision processes (MDPs). Given a set of expert policies trained on a state and actio…
cs.LG2016★ 5 cited
Recursive Regression with Neural Networks: Approximating the HJI PDE Solution
Vicenç Rubies-Royo, Claire Tomlin
The majority of methods used to compute approximations to the Hamilton-Jacobi-Isaacs partial differential equation (HJI PDE) rely on the discretization of the state space to perfor…