35 citations · 89 across the 3 of their papers we have counts for
3 papers
cs.AI2013★ 35 cited
Decision-Theoretic Planning with Concurrent Temporally Extended Actions
Khashayar Rohanimanesh, Sridhar Mahadevan
We investigate a model for planning under uncertainty with temporallyextended actions, where multiple actions can be taken concurrently at each decision epoch. Our model is based o…
cs.LG2012★ 21 cited
Sparse Q-learning with Mirror Descent
Sridhar Mahadevan, Bo Liu
This paper explores a new framework for reinforcement learning based on online convex optimization, in particular mirror descent and related algorithms. Mirror descent can be viewe…
cs.AI2012★ 33 cited
Representation Policy Iteration
Sridhar Mahadevan
This paper addresses a fundamental issue central to approximation methods for solving large Markov decision processes (MDPs): how to automatically learn the underlying representati…