27 citations · 27 across the 1 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2023
Monte-Carlo Search for an Equilibrium in Dec-POMDPs
Yang You, Vincent Thomas, Francis Colas +1
Decentralized partially observable Markov decision processes (Dec-POMDPs) formalize the problem of designing individual controllers for a group of collaborative agents under stocha…
cs.AI2012★ 27 cited
Near-Optimal BRL using Optimistic Local Transitions
Mauricio Araya, Olivier Buffet, Vincent Thomas
Model-based Bayesian Reinforcement Learning (BRL) allows a found formalization of the problem of acting optimally while facing an unknown environment, i.e., avoiding the exploratio…