BCQQ: Batch-Constraint Quantum Q-Learning with Cyclic Data Re-uploading
arXiv:2305.00905 · doi:10.1109/IJCNN60899.2024.10651268
Abstract
Deep reinforcement learning (DRL) often requires a large number of data and environment interactions, making the training process time-consuming. This challenge is further exacerbated in the case of batch RL, where the agent is trained solely on a pre-collected dataset without environment interactions. Recent advancements in quantum computing suggest that quantum models might require less data for training compared to classical methods. In this paper, we investigate this potential advantage by proposing a batch RL algorithm that utilizes VQC as function approximators within the discrete batch-constraint deep Q-learning (BCQ) algorithm. Additionally, we introduce a novel data re-uploading scheme by cyclically shifting the order of input variables in the data encoding layers. We evaluate the efficiency of our algorithm on the OpenAI CartPole environment and compare its performance to the classical neural network-based discrete BCQ.
References in corpus (9)
- On the Convergence of Adam and Beyond
- Quantum reinforcement learning
- Benchmarking Batch Deep Reinforcement Learning Algorithms
- A Comparison of Various Classical Optimizers for a Variational Quantum Linear Solver
- Quantum computing of the Li nucleus via ordered unitary coupled clusters
- A Survey on Quantum Reinforcement Learning
- An Empirical Comparison of Optimizers for Quantum Machine Learning with SPSA-based Gradients
- Hybrid quantum-classical classifier based on tensor network and variational quantum circuit
- A scale-dependent notion of effective dimension