1 citations · 1 across the 4 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2022
Tight Performance Guarantees of Imitator Policies with Continuous Actions
Davide Maran, Alberto Maria Metelli, Marcello Restelli
Behavioral Cloning (BC) aims at learning a policy that mimics the behavior demonstrated by an expert. The current theoretical understanding of BC is limited to the case of finite a…
cs.LG2022★ 1 cited
Delayed Reinforcement Learning by Imitation
Pierre Liotet, Davide Maran, Lorenzo Bisi +1
When the agent's observations or interactions are delayed, classic reinforcement learning tools usually fail. In this paper, we propose a simple yet new and efficient solution to t…
cs.LG2021
A note on the article "On Exploiting Spectral Properties for Solving MDP with Large State Space"
D. Maran
We improve a theoretical result of the article "On Exploiting Spectral Properties for Solving MDP with Large State Space" showing that their algorithm, which was proved to converge…