Adversarial Deep Reinforcement Learning in Portfolio Management
arXiv:1808.09940
Abstract
In this paper, we implement three state-of-art continuous reinforcement learning algorithms, Deep Deterministic Policy Gradient (DDPG), Proximal Policy Optimization (PPO) and Policy Gradient (PG)in portfolio management. All of them are widely-used in game playing and robot control. What's more, PPO has appealing theoretical propeties which is hopefully potential in portfolio management. We present the performances of them under different settings, including different learning rates, objective functions, feature combinations, in order to provide insights for parameters tuning, features selection and data preparation. We also conduct intensive experiments in China Stock market and show that PG is more desirable in financial market than DDPG and PPO, although both of them are more advanced. What's more, we propose a so called Adversarial Training method and show that it can greatly improve the training efficiency and significantly promote average daily return and sharpe ratio in back test. Based on this new modification, our experiments results show that our agent based on Policy Gradient can outperform UCRP.
Cited by in corpus (11)
- Model-based Deep Reinforcement Learning for Dynamic Portfolio Optimization
- Deep Reinforcement Learning for Long-Short Portfolio Optimization
- Model-Free Reinforcement Learning for Financial Portfolios: A Brief Survey
- Capturing Financial markets to apply Deep Reinforcement Learning
- Deep Reinforcement Learning in Quantitative Algorithmic Trading: A Review
- Reinforcement Learning for Quantitative Trading
- Deep Stock Trading: A Hierarchical Reinforcement Learning Framework for Portfolio Optimization and Order Execution
- Reinforcement-Learning based Portfolio Management with Augmented Asset Movement Prediction States
- Deep Reinforcement Learning for Portfolio Optimization using Latent Feature State Space (LFSS) Module
- A General Framework on Enhancing Portfolio Management with Reinforcement Learning
- A Deep Deterministic Policy Gradient-based Strategy for Stocks Portfolio Management