From the 1 of 9 linked papers with an AI index.
3 papers · 1 filter
Drift Q-Learning
Anas Houssaini, Mohamad H. Danesh, Amin Abyaneh +3
Offline reinforcement learning requires improving a policy from fixed data while avoiding out-of-distribution actions with unreliable value estimates. Diffusion and flow policies h…
Contractive Diffusion Policies: Robust Action Diffusion via Contractive Score-Based Sampling with Differential Equations
Amin Abyaneh, Charlotte Morissette, Mohamad H. Danesh +4
Diffusion policies have emerged as powerful generative models for offline policy learning, whose sampling process can be rigorously characterized by a score function guiding a stoc…
Contractive Dynamical Imitation Policies for Efficient Out-of-Sample Recovery
Amin Abyaneh, Mahrokh G. Boroujeni, Hsiu-Chin Lin +1
Imitation learning is a data-driven approach to learning policies from expert behavior, but it is prone to unreliable outcomes in out-of-sample (OOS) regions. While previous resear…