3 papers
cs.LG2024
Fair Resource Allocation in Weakly Coupled Markov Decision Processes
Xiaohui Tu, Yossiri Adulyasak, Nima Akbarzadeh +1
We consider fair resource allocation in sequential decision-making environments modeled as weakly coupled Markov decision processes, where resource constraints couple the action sp…
cs.LG2024
Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
Nima Akbarzadeh, Yossiri Adulyasak, Erick Delage
In restless bandits, a central agent is tasked with optimally distributing limited resources across several bandits (arms), with each arm being a Markov decision process. In this w…
cs.LG2023
Approximate information state based convergence analysis of recurrent Q-learning
Erfan Seyedsalehi, Nima Akbarzadeh, Amit Sinha +1
In spite of the large literature on reinforcement learning (RL) algorithms for partially observable Markov decision processes (POMDPs), a complete theoretical understanding is stil…