2 papers
cs.LG2026
Value Bonuses using Ensemble Errors for Exploration in Reinforcement Learning
Abdul Wahab, Raksha Kumaraswamy, Martha White
Optimistic value estimates provide one mechanism for directed exploration in reinforcement learning (RL). The agent acts greedily with respect to an estimate of the value plus what…
eess.SY2024
PC-Gym: Benchmark Environments For Process Control Problems
Maximilian Bloor, José Torraca, Ilya Orson Sandoval +6
PC-Gym is an open-source tool for developing and evaluating reinforcement learning (RL) algorithms in chemical process control. It features environments that simulate various chemi…