2 papers
cs.LG2026
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
Juan Sebastian Rojas, Chi-Guhn Lee
The temporal difference (TD) error was first formalized in Sutton (1988), where it was first characterized as the difference between temporally successive predictions, and later, i…
cs.MA2025
Hypernetwork-based approach for optimal composition design in partially controlled multi-agent systems
Kyeonghyeon Park, David Molina Concha, Hyun-Rok Lee +2
Partially Controlled Multi-Agent Systems (PCMAS) are comprised of controllable agents, managed by a system designer, and uncontrollable agents, operating autonomously. This study a…