3 papers
cs.MA2025
Predictive Auxiliary Learning for Belief-based Multi-Agent Systems
Qinwei Huang, Stefan Wang, Simon Khan +2
The performance of multi-agent reinforcement learning (MARL) in partially observable environments depends on effectively aggregating information from observations, communications,…
cs.DS2025
Linearithmic Clean-up for Vector-Symbolic Key-Value Memory with Kroneker Rotation Products
Ruipeng Liu, Qinru Qiu, Simon Khan +1
A computational bottleneck in current Vector-Symbolic Architectures (VSAs) is the ``clean-up'' step, which decodes the noisy vectors retrieved from the architecture. Clean-up typic…
cs.AI2024
Why the Agent Made that Decision: Contrastive Explanation Learning for Reinforcement Learning
Rui Zuo, Simon Khan, Zifan Wang +2
Reinforcement learning (RL) has demonstrated remarkable success in solving complex decision-making problems, yet its adoption in critical domains is hindered by the lack of interpr…