Showing stat.MLShow all
2 papers · 1 filter
stat.ML2025
Efficient Risk-sensitive Planning via Entropic Risk Measures
Alexandre Marthe, Samuel Bounan, Aurélien Garivier +1
Risk-sensitive planning aims to identify policies maximizing some tail-focused metrics in Markov Decision Processes (MDPs). Such an optimization task can be very costly for the mos…
stat.ML2025
Sequential Learning of the Pareto Front for Multi-objective Bandits
Elise Crépon, Aurélien Garivier, Wouter M Koolen
We study the problem of sequential learning of the Pareto front in multi-objective multi-armed bandits. An agent is faced with K possible arms to pull. At each turn she picks one,…