10 citations · 22 across the 7 of their papers we have counts for
4 papers · 1 filter
Multiagent Rollout and Policy Iteration for POMDP with Application to Multi-Robot Repair Problems
Sushmita Bhattacharya, Siva Kailas, Sahil Badyal +2
In this paper we consider infinite horizon discounted dynamic programming problems with finite state and control spaces, partial state observations, and a multiagent structure. We…
Multiagent Value Iteration Algorithms in Dynamic Programming and Reinforcement Learning
Dimitri Bertsekas
We consider infinite horizon dynamic programming problems, where the control at each stage consists of several distinct decisions, each one made by one of several agents. In an ear…
Reinforcement Learning for POMDP: Partitioned Rollout and Policy Iteration with Application to Autonomous Sequential Repair Problems
Sushmita Bhattacharya, Sahil Badyal, Thomas Wheeler +2
In this paper we consider infinite horizon discounted dynamic programming problems with finite state and control spaces, and partial state observations. We discuss an algorithm tha…
Constrained Multiagent Rollout and Multidimensional Assignment with the Auction Algorithm
Dimitri Bertsekas
We consider an extension of the rollout algorithm that applies to constrained deterministic dynamic programming, including challenging combinatorial optimization problems. The algo…