2 papers
cs.LG2024
Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning Approach
Johan Peralez, Aurèlien Delage, Jacopo Castellini +2
The centralized training for decentralized execution paradigm emerged as the state-of-the-art approach to -optimally solving decentralized partially observable Markov decision p…
cs.MA2023
On Convex Optimal Value Functions For POSGs
Rafael F. Cunha, Jacopo Castellini, Johan Peralez +1
Multi-agent planning and reinforcement learning can be challenging when agents cannot see the state of the world or communicate with each other due to communication costs, latency,…