Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Efficient Matroid Bandit Linear Optimization Leveraging Unimodality
Aurélien Delage, Romaric Gaudel
We study the combinatorial semi-bandit problem under matroid constraints. The regret achieved by recent approaches is optimal, in the sense that it matches the lower bound. Yet, ti…
cs.LG2025
Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning Approach
Johan Peralez, Aurèlien Delage, Jacopo Castellini +2
The centralized training for decentralized execution paradigm emerged as the state-of-the-art approach to -optimally solving decentralized partially observable Markov decision…