2 papers
cs.LG2025
Graph of Agents: Principled Long Context Modeling by Emergent Multi-Agent Collaboration
Taejong Joo, Shu Ishida, Ivan Sosnovik +4
As a model-agnostic approach to long context modeling, multi-agent systems can process inputs longer than a large language model's context window without retraining or architectura…
cs.LG2024
Fairness in Reinforcement Learning with Bisimulation Metrics
Sahand Rezaei-Shoshtari, Hanna Yurchyk, Scott Fujimoto +2
Ensuring long-term fairness is crucial when developing automated decision making systems, specifically in dynamic and sequential environments. By maximizing their reward without co…