A Benchmark Environment Motivated by Industrial Control Problems
arXiv:1709.09480 · doi:10.1109/SSCI.2017.8280935
Abstract
In the research area of reinforcement learning (RL), frequently novel and promising methods are developed and introduced to the RL community. However, although many researchers are keen to apply their methods on real-world problems, implementing such methods in real industry environments often is a frustrating and tedious process. Generally, academic research groups have only limited access to real industrial data and applications. For this reason, new methods are usually developed, evaluated and compared by using artificial software benchmarks. On one hand, these benchmarks are designed to provide interpretable RL training scenarios and detailed insight into the learning process of the method on hand. On the other hand, they usually do not share much similarity with industrial real-world applications. For this reason we used our industry experience to design a benchmark which bridges the gap between freely available, documented, and motivated artificial benchmarks and properties of real industrial problems. The resulting industrial benchmark (IB) has been made publicly available to the RL community by publishing its Java and Python code, including an OpenAI Gym wrapper, on Github. In this paper we motivate and describe in detail the IB's dynamics and identify prototypic experimental settings that capture common situations in real-world industry control problems.
References in corpus (3)
Cited by in corpus (13)
- Deep Reinforcement Learning: An Overview
- LIFT: Reinforcement Learning in Computer Systems by Learning From Demonstrations
- NeoRL: A Near Real-World Benchmark for Offline Reinforcement Learning
- Generating Interpretable Fuzzy Controllers using Particle Swarm Optimization and Genetic Programming
- Off-Policy Reinforcement Learning with Delayed Rewards
- Learning Long-Term Reward Redistribution via Randomized Return Decomposition
- Learning Guidance Rewards with Trajectory-space Smoothing
- Learning Control Policies for Variable Objectives from Offline Data
- Concept and the implementation of a tool to convert industry 4.0 environments modeled as FSM to an OpenAI Gym wrapper
- Bellman: A Toolbox for Model-Based Reinforcement Learning in TensorFlow
- Showing Your Offline Reinforcement Learning Work: Online Evaluation Budget Matters
- Wield: Systematic Reinforcement Learning With Progressive Randomization
- Augmenting PID Control with Deep Reinforcement Learning: A Hybrid Approach to the Industrial Benchmark