Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Accelerating Policy Synthesis in Large-Scale MDPs via Hierarchical Adaptive Refinement
Alexandros Evangelidis, Gricel Vázquez, Simos Gerasimou
Software-intensive systems, such as software product lines and robotics, utilise Markov decision processes (MDPs) to capture uncertainty and analyse sequential decision-making prob…
cs.AI2026
Mind the Prompt: Self-adaptive Generation of Task Plan Explanations via LLMs
Gricel Vázquez, Alexandros Evangelidis, Sepeedeh Shahbeigi +2
Integrating Large Language Models (LLMs) into complex software systems enables the generation of human-understandable explanations of opaque AI processes, such as automated task pl…
cs.AI2025
Safe Reinforcement Learning in Black-Box Environments via Adaptive Shielding
Daniel Bethell, Simos Gerasimou, Radu Calinescu +1
Empowering safe exploration of reinforcement learning (RL) agents during training is a critical challenge towards their deployment in many real-world scenarios. When prior knowledg…