Multi-Agent Deep Reinforcement Learning for Dynamic Power Allocation in Wireless Networks
arXiv:1808.00490 · doi:10.1109/JSAC.2019.2933973
Abstract
This work demonstrates the potential of deep reinforcement learning techniques for transmit power control in wireless networks. Existing techniques typically find near-optimal power allocations by solving a challenging optimization problem. Most of these algorithms are not scalable to large networks in real-world scenarios because of their computational complexity and instantaneous cross-cell channel state information (CSI) requirement. In this paper, a distributively executed dynamic power allocation scheme is developed based on model-free deep reinforcement learning. Each transmitter collects CSI and quality of service (QoS) information from several neighbors and adapts its own transmit power accordingly. The objective is to maximize a weighted sum-rate utility function, which can be particularized to achieve maximum sum-rate or proportionally fair scheduling. Both random variations and delays in the CSI are inherently addressed using deep Q-learning. For a typical network architecture, the proposed algorithm is shown to achieve near-optimal power allocation in real time based on delayed CSI measurements available to the agents. The proposed scheme is especially suitable for practical scenarios where the system model is inaccurate and CSI delay is non-negligible.
12 pages, 7 figures, submitted. v2: the updated title, in addition to improved readability. v3: revised
References in corpus (2)
Cited by in corpus (30)
- Edge Learning for B5G Networks with Distributed Signal Processing: Semantic Communication, Edge Computing, and Wireless Sensing
- Generative AI Agents with Large Language Model for Satellite Networks via a Mixture of Experts Transmission
- Resource Management in Wireless Networks via Multi-Agent Deep Reinforcement Learning
- Enabling AI in Future Wireless Networks: A Data Life Cycle Perspective
- Learning-Based Downlink Power Allocation in Cell-Free Massive MIMO Systems
- Network Slicing with MEC and Deep Reinforcement Learning for the Internet of Vehicles
- Learning Decentralized Wireless Resource Allocations with Graph Neural Networks
- Deep Reinforcement Learning for Radio Resource Allocation and Management in Next Generation Heterogeneous Wireless Networks: A Survey
- Deep Unsupervised Learning for Joint Antenna Selection and Hybrid Beamforming
- Learning to Branch: Accelerating Resource Allocation in Wireless Networks
- Link Scheduling using Graph Neural Networks
- Learning to Continuously Optimize Wireless Resource in a Dynamic Environment: A Bilevel Optimization Perspective
- Multi-Agent Deep Reinforcement Learning for Distributed Resource Management in Wirelessly Powered Communication Networks
- State-Augmented Learnable Algorithms for Resource Management in Wireless Networks
- Network slicing for vehicular communications: a multi-agent deep reinforcement learning approach
- Cellular Network Capacity and Coverage Enhancement with MDT Data and Deep Reinforcement Learning
- Hybrid Beamforming for mmWave MU-MISO Systems Exploiting Multi-agent Deep Reinforcement Learning
- Model-Free Learning of Optimal Ergodic Policies in Wireless Systems
- A Context-aware Radio Resource Management in Heterogeneous Virtual RANs
- Effective Multi-User Delay-Constrained Scheduling with Deep Recurrent Reinforcement Learning
- Conservative and Risk-Aware Offline Multi-Agent Reinforcement Learning
- Learning to Transmit with Provable Guarantees in Wireless Federated Learning
- Percentile Optimization in Wireless Networks- Part II: Beamforming for Cell-Edge Throughput Maximization
- A Multi-Agent Reinforcement Learning Framework for Public Health Decision Analysis
- Asynchronous Message-Passing and Zeroth-Order Optimization Based Distributed Learning with a Use-Case in Resource Allocation in Communication Networks
- Asymptotic Performance of Box-RLS Decoders under Imperfect CSI with Optimized Resource Allocation
- A Deep Learning Based Resource Allocator for Communication Networks with Dynamic User Utility Demands
- Fast-Fading Channel and Power Optimization of the Magnetic Inductive Cellular Network
- Minimizing File Transfer Time in Opportunistic Spectrum Access Model
- Optimal wireless rate and power control in the presence of jammers using reinforcement learning