3 papers
cs.LG2025
SACn: Soft Actor-Critic with n-step Returns
Jakub Łyskawa, Jakub Lewandowski, Paweł Wawrzyński
Soft Actor-Critic (SAC) is widely used in practical applications and is now one of the most relevant off-policy online model-free reinforcement learning (RL) methods. The technique…
cs.LG2022
Emergency action termination for immediate reaction in hierarchical reinforcement learning
Michał Bortkiewicz, Jakub Łyskawa, Paweł Wawrzyński +3
Hierarchical decomposition of control is unavoidable in large dynamical systems. In reinforcement learning (RL), it is usually solved with subgoals defined at higher policy levels…
cs.LG2020
A framework for reinforcement learning with autocorrelated actions
Marcin Szulc, Jakub Łyskawa, Paweł Wawrzyński
The subject of this paper is reinforcement learning. Policies are considered here that produce actions based on states and random elements autocorrelated in subsequent time instant…