1 citations · 1 across the 3 of their papers we have counts for
4 papers
Tight Performance Guarantees of Imitator Policies with Continuous Actions
Davide Maran, Alberto Maria Metelli, Marcello Restelli
Behavioral Cloning (BC) aims at learning a policy that mimics the behavior demonstrated by an expert. The current theoretical understanding of BC is limited to the case of finite a…
Delayed Reinforcement Learning by Imitation
Pierre Liotet, Davide Maran, Lorenzo Bisi +1
When the agent's observations or interactions are delayed, classic reinforcement learning tools usually fail. In this paper, we propose a simple yet new and efficient solution to t…
A note on the article "On Exploiting Spectral Properties for Solving MDP with Large State Space"
D. Maran
We improve a theoretical result of the article "On Exploiting Spectral Properties for Solving MDP with Large State Space" showing that their algorithm, which was proved to converge…
Least singular value and condition number of a square random matrix with i.i.d. rows
Matteo Gregoratti, Davide Maran
We consider a square random matrix made by i.i.d. rows with any distribution and prove that, for any given dimension, the probability for the least singular value to be in [0; )…