39 citations · 57 across the 17 of their papers we have counts for
1 paper · 1 filter
Mohammadhosein Hasanbeig, Alessandro Abate, Daniel Kroening
We propose a method for efficient training of Q-functions for continuous-state Markov Decision Processes (MDPs) such that the traces of the resulting policies satisfy a given Linea…