Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Online MDP with Transition Prototypes: A Robust Adaptive Approach
Shuo Sun, Meng Qi, Zuo-Jun Max Shen
In this work, we consider an online robust Markov Decision Process (MDP) where we have the information of finitely many prototypes of the underlying transition kernel. We consider…
cs.LG2021
Smart Feasibility Pump: Reinforcement Learning for (Mixed) Integer Programming
Meng Qi, Mengxin Wang, Zuo-Jun Shen
In this work, we propose a deep reinforcement learning (DRL) model for finding a feasible solution for (mixed) integer programming (MIP) problems. Finding a feasible solution for M…