45 citations · 168 across the 23 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.LG2024
Operator World Models for Reinforcement Learning
Pietro Novelli, Marco Pratticò, Massimiliano Pontil +1
Policy Mirror Descent (PMD) is a powerful and theoretically sound methodology for sequential decision-making. However, it is not directly applicable to Reinforcement Learning (RL)…
stat.ML2024
Closed-form Filtering for Non-linear Systems
Théophile Cantelobre, Carlo Ciliberto, Benjamin Guedj +1
Sequential Bayesian Filtering aims to estimate the current state distribution of a Hidden Markov Model, given the past observations. The problem is well-known to be intractable for…