1 paper · 1 filter
Denis Tarasov, Robert K. Katzschmann
Recent offline reinforcement learning methods increasingly rely on expressive generative policies and specialized value-guidance mechanisms. We ask whether comparable progress can…