2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.AI2024
Simulation-Based Optimistic Policy Iteration For Multi-Agent MDPs with Kullback-Leibler Control Cost
Khaled Nakhleh, Ceyhun Eksin, Sabit Ekin
This paper proposes an agent-based optimistic policy iteration (OPI) scheme for learning stationary optimal stochastic policies in multi-agent Markov Decision Processes (MDPs), in…
cs.LG2022★ 2 cited
DeepTOP: Deep Threshold-Optimal Policy for MDPs and RMABs
Khaled Nakhleh, I-Hong Hou
We consider the problem of learning the optimal threshold policy for control problems. Threshold policies make control decisions by evaluating whether an element of the system stat…
cs.NI2022
A Theory of Second-Order Wireless Network Optimization and Its Application on AoI
Daojing Guo, Khaled Nakhleh, I-Hong Hou +2
This paper introduces a new theoretical framework for optimizing second-order behaviors of wireless networks. Unlike existing techniques for network utility maximization, which onl…