3 papers
cs.LO2026
A Theory of Hanoi Omega-Automata and Games
Emmanuel Filiot, Allen Joseph, Guillermo A. Pérez +1
The Hanoi Omega-Automata (HOA) format has established itself as the definitive standard for encoding -regular automata in modern synthesis tools. While HOA is widely adopted due…
cs.AI2026
Computing the Reachability Value of Posterior-Deterministic POMDPs
Nathanaël Fijalkow, Arka Ghosh, Roman Kniazev +2
Partially observable Markov decision processes (POMDPs) are a fundamental model for sequential decision-making under uncertainty. However, many verification and synthesis problems…
cs.AI2025
Data-Efficient Safe Policy Improvement Using Parametric Structure
Kasper Engelen, Guillermo A. Pérez, Marnix Suilen
Safe policy improvement (SPI) is an offline reinforcement learning problem in which a new policy that reliably outperforms the behavior policy with high confidence needs to be comp…