3 papers
cs.LG2025
Implicit Repair with Reinforcement Learning in Emergent Communication
Fábio Vital, Alberto Sardinha, Francisco S. Melo
Conversational repair is a mechanism used to detect and resolve miscommunication and misinformation problems when two or more agents interact. One particular and underexplored form…
cs.LG2025
Distributed Value Decomposition Networks with Networked Agents
Guilherme S. Varela, Alberto Sardinha, Francisco S. Melo
We investigate the problem of distributed training under partial observability, whereby cooperative multi-agent reinforcement learning agents (MARL) maximize the expected cumulativ…
cs.LG2025
Networked Agents in the Dark: Team Value Learning under Partial Observability
Guilherme S. Varela, Alberto Sardinha, Francisco S. Melo
We propose a novel cooperative multi-agent reinforcement learning (MARL) approach for networked agents. In contrast to previous methods that rely on complete state information or j…