1 citations · 2 across the 9 of their papers we have counts for
4 papers · 1 filter
Applying Policy Iteration for Training Recurrent Neural Networks
I. Szita, A. Lorincz
Recurrent neural networks are often used for learning time-series data. Based on a few assumptions we model this learning task as a minimization problem of a nonlinear least-square…
Kalman-filtering using local interactions
Barnabas Poczos, Andras Lorincz
There is a growing interest in using Kalman-filter models for brain modelling. In turn, it is of considerable importance to represent Kalman-filter in connectionist forms with loca…
Temporal plannability by variance of the episode length
Balint Takacs, Istvan Szita, Andras Lorincz
Optimization of decision problems in stochastic environments is usually concerned with maximizing the probability of achieving the goal and minimizing the expected episode length.…
Searching for Plannable Domains can Speed up Reinforcement Learning
Istvan Szita, Balint Takacs, Andras Lorincz
Reinforcement learning (RL) involves sequential decision making in uncertain environments. The aim of the decision-making agent is to maximize the benefit of acting in its environm…