3 papers
cs.LG2025
Distributed Value Decomposition Networks with Networked Agents
Guilherme S. Varela, Alberto Sardinha, Francisco S. Melo
We investigate the problem of distributed training under partial observability, whereby cooperative multi-agent reinforcement learning agents (MARL) maximize the expected cumulativ…
cs.LG2025
Networked Agents in the Dark: Team Value Learning under Partial Observability
Guilherme S. Varela, Alberto Sardinha, Francisco S. Melo
We propose a novel cooperative multi-agent reinforcement learning (MARL) approach for networked agents. In contrast to previous methods that rely on complete state information or j…
eess.SY2021
A Methodology for the Development of RL-Based Adaptive Traffic Signal Controllers
Guilherme S. Varela, Pedro P. Santos, Alberto Sardinha +1
This article proposes a methodology for the development of adaptive traffic signal controllers using reinforcement learning. Our methodology addresses the lack of standardization i…