Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
Stefan Stojanovic, Alexandre Proutiere
Hierarchical reinforcement learning can improve generalization by decomposing long-horizon decision-making into simpler subproblems. However, existing approaches often rely on rest…
cs.LG2025
Adaptive Reinforcement Learning for Unobservable Random Delays
John Wikman, Alexandre Proutiere, David Broman
In standard reinforcement learning (RL) settings, the interaction between the agent and the environment is typically modeled as a Markov decision process (MDP), which assumes that…