1 citations · 1 across the 3 of their papers we have counts for
4 papers
Formally-Sharp DAgger for MCTS: Lower-Latency Monte Carlo Tree Search using Data Aggregation with Formal Methods
Debraj Chakraborty, Damien Busatto-Gaston, Jean-François Raskin +1
We study how to efficiently combine formal methods, Monte Carlo Tree Search (MCTS), and deep learning in order to produce high-quality receding horizon policies in large Markov Dec…
Bi-Objective Lexicographic Optimization in Markov Decision Processes with Related Objectives
Damien Busatto-Gaston, Debraj Chakraborty, Anirban Majumdar +3
We consider lexicographic bi-objective problems on Markov Decision Processes (MDPs), where we optimize one objective while guaranteeing optimality of another. We propose a two-stag…
Strategy Synthesis for Global Window PCTL
Benjamin Bordais, Damien Busatto-Gaston, Shibashis Guha +1
Given a Markov decision process (MDP) and a formula , the strategy synthesis problem asks if there exists a strategy s.t. the resulting Markov chain satisfies …
Monte Carlo Tree Search guided by Symbolic Advice for MDPs
Damien Busatto-Gaston, Debraj Chakraborty, Jean-Francois Raskin
In this paper, we consider the online computation of a strategy that aims at optimizing the expected average reward in a Markov decision process. The strategy is computed with a re…