1 paper · 1 filter
Philippe Hansen-Estruch, Ilya Kostrikov, Michael Janner +2
Effective offline RL methods require properly handling out-of-distribution actions. Implicit Q-learning (IQL) addresses this by training a Q-function using only dataset actions thr…