2 citations · 4 across the 3 of their papers we have counts for
1 paper · 1 filter
Matthew Chang, Arjun Gupta, Saurabh Gupta
This paper tackles the problem of learning value functions from undirected state-only experience (state transitions without action labels i.e. (s,s',r) tuples). We first theoretica…