Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
PADME: Procedure Aware DynaMic Execution
Deepeka Garg, Sihan Zeng, Annapoorani L. Narayanan +2
Learning to autonomously execute long-horizon procedures from natural language remains a core challenge for intelligent agents. Free-form instructions such as recipes, scientific p…
cs.AI2025
Learning Robust Reward Machines from Noisy Labels
Roko Parac, Lorenzo Nodari, Leo Ardon +3
This paper presents PROB-IRM, an approach that learns robust reward machines (RMs) for reinforcement learning (RL) agents from noisy execution traces. The key aspect of RM-driven R…
cs.AI2025
FORM: Learning Expressive and Transferable First-Order Logic Reward Machines
Leo Ardon, Daniel Furelos-Blanco, Roko Parac +1
Reward machines (RMs) are an effective approach for addressing non-Markovian rewards in reinforcement learning (RL) through finite-state machines. Traditional RMs, which label edge…