Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Actor-Critic with Active Importance Sampling
Majid Molaei, Gabor Paczolay, Matteo Papini +2
This paper introduces the Active-Importance-Sampling Actor-Critic (AISAC) algorithm, an extension of the Actor-Critic framework for reducing variance in policy gradient estimation.…
cs.LG2024
Statistical Analysis of Policy Space Compression Problem
Majid Molaei, Marcello Restelli, Alberto Maria Metelli +1
Policy search methods are crucial in reinforcement learning, offering a framework to address continuous state-action and partially observable problems. However, the complexity of e…