Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Minimax PAC Bounds for Learning in Exogenous Contextual MDPs
Corentin Pla, Hugo Richard, Marc Abeille +1
We study PAC learning in tabular discounted Markov decision processes with exogenous i.i.d. contexts, with discount factor , finite state space , action space $\mat…
stat.ML2026
On the Hardness of Reinforcement Learning with Transition Look-Ahead
Corentin Pla, Hugo Richard, Marc Abeille +2
We study reinforcement learning (RL) with transition look-ahead, where the agent may observe which states would be visited upon playing any sequence of actions before decidi…