455 citations · 971 across the 43 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2018
Efficient Exploration through Bayesian Deep Q-Networks
Kamyar Azizzadenesheli, Animashree Anandkumar
We study reinforcement learning (RL) in high dimensional episodic Markov decision processes (MDP). We consider value-based RL when the optimal Q-value is a linear function of d-dim…
cs.AI2017
Experimental results : Reinforcement Learning of POMDPs using Spectral Methods
Kamyar Azizzadenesheli, Alessandro Lazaric, Animashree Anandkumar
We propose a new reinforcement learning algorithm for partially observable Markov decision processes (POMDP) based on spectral decomposition methods. While spectral methods have be…