3 papers
cs.LG2025
Efficient Matroid Bandit Linear Optimization Leveraging Unimodality
Aurélien Delage, Romaric Gaudel
We study the combinatorial semi-bandit problem under matroid constraints. The regret achieved by recent approaches is optimal, in the sense that it matches the lower bound. Yet, ti…
cs.LG2025
Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning Approach
Johan Peralez, Aurèlien Delage, Jacopo Castellini +2
The centralized training for decentralized execution paradigm emerged as the state-of-the-art approach to -optimally solving decentralized partially observable Markov decision…
cs.GT2025
Solving Hierarchical Information-Sharing Dec-POMDPs: An Extensive-Form Game Approach
Johan Peralez, Aurélien Delage, Olivier Buffet +1
A recent theory shows that a multi-player decentralized partially observable Markov decision process can be transformed into an equivalent single-player game, enabling the applicat…