17 citations · 33 across the 7 of their papers we have counts for
Showing 2003Show all
3 papers · 1 filter
cs.LG2003
Reinforcement Learning with Linear Function Approximation and LQ control Converges
Istvan Szita, Andras Lorincz
Reinforcement learning is commonly used with function approximation. However, very few positive results are known about the convergence of function approximation based RL control a…
cs.LG2003
Kalman filter control in the reinforcement learning framework
Istvan Szita, Andras Lorincz
There is a growing interest in using Kalman-filter models in brain modelling. In turn, it is of considerable importance to make Kalman-filters amenable for reinforcement learning.…
cs.AI2003
Temporal plannability by variance of the episode length
Balint Takacs, Istvan Szita, Andras Lorincz
Optimization of decision problems in stochastic environments is usually concerned with maximizing the probability of achieving the goal and minimizing the expected episode length.…