2 papers
cs.LG2026
Risk-Aware General-Utility Markov Decision Processes
Pedro P. Santos, Fábio Vital, Alberto Sardinha +1
We study general-utility Markov decision processes (GUMDPs) with risk-aware objectives. In this framework, an agent aims to optimize a risk measure of the distribution of objective…
cs.LG2025
Implicit Repair with Reinforcement Learning in Emergent Communication
Fábio Vital, Alberto Sardinha, Francisco S. Melo
Conversational repair is a mechanism used to detect and resolve miscommunication and misinformation problems when two or more agents interact. One particular and underexplored form…