19 citations · 19 across the 1 of their papers we have counts for
1 paper
M. M. Hassan Mahmud, Majd Hawasly, Benjamin Rosman +1
We present algorithms to effectively represent a set of Markov decision processes (MDPs), whose optimal policies have already been learned, by a smaller source subset for lifelong,…