49 citations · 87 across the 15 of their papers we have counts for
4 papers · 1 filter
The many faces of optimism - Extended version
István Szita, András Lőrincz
The exploration-exploitation dilemma has been an intriguing and unsolved problem within the framework of reinforcement learning. "Optimism in the face of uncertainty" and model bui…
Online variants of the cross-entropy method
Istvan Szita, Andras Lorincz
The cross-entropy method is a simple but efficient method for global optimization. In this paper we provide two online variants of the basic CEM, together with a proof of convergen…
D-optimal Bayesian Interrogation for Parameter and Noise Identification of Recurrent Neural Networks
Barnabas Poczos, Andras Lorincz
We introduce a novel online Bayesian method for the identification of a family of noisy recurrent neural networks (RNNs). We develop Bayesian active learning technique in order to…
Factored Value Iteration Converges
Istvan Szita, Andras Lorincz
In this paper we propose a novel algorithm, factored value iteration (FVI), for the approximate solution of factored Markov decision processes (fMDPs). The traditional approximate…