18 citations · 18 across the 3 of their papers we have counts for
3 papers
Controlled Descent Training
Viktor Andersson, Balázs Varga, Vincent Szolnoky +3
In this work, a novel and model-based artificial neural network (ANN) training method is developed supported by optimal control theory. The method augments training labels in order…
Deep Q-learning: a robust control approach
Balazs Varga, Balazs Kulcsar, Morteza Haghir Chehreghani
In this paper, we place deep Q-learning into a control-oriented perspective and study its learning dynamics with well-established techniques from robust control. We formulate an un…
Constrained Policy Gradient Method for Safe and Fast Reinforcement Learning: a Neural Tangent Kernel Based Approach
Balázs Varga, Balázs Kulcsár, Morteza Haghir Chehreghani
This paper presents a constrained policy gradient algorithm. We introduce constraints for safe learning with the following steps. First, learning is slowed down (lazy learning) so…