2 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Fan Lu, Sean Meyn
The paper introduces the first formulation of convex Q-learning for Markov decision processes with function approximation. The algorithms and theory rest on a relaxation of a dual…