activity
20212026
most citedEmpirical analysis of PGA-MAP-Elites for Neuroevolution in Uncertain Domains

21 citations · 47 across the 14 of their papers we have counts for

collaborators

14 papers

cs.LG2026

Out-of-Distribution Generalisation with Sequence Models in Offline Multi-Agent Reinforcement Learning

Oussama Hidaoui, Omer Ebead, Ulrich Armel Mbou Sob +14

Generalising to unseen tasks remains a fundamental challenge in offline multi-agent reinforcement learning (MARL). In this work, we present a principled analysis of zero-shot task…

cs.LG2026

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation

Asim Osman, Sasha Abramowitz, Mark Bergh +13

Contrastive reinforcement learning (CRL) learns goal-conditioned Q-values through a contrastive objective over state-action and goal representations, removing the need for hand-cra…

cs.LG2025

Oryx: a Scalable Sequence Model for Many-Agent Coordination in Offline MARL

Claude Formanek, Omayma Mahjoub, Louay Ben Nessir +10

A key challenge in offline multi-agent reinforcement learning (MARL) is achieving effective many-agent multi-step coordination in complex environments. In this work, we propose Ory…

cs.LG2025

Breaking the Performance Ceiling in Reinforcement Learning requires Inference Strategies

Felix Chalumeau, Daniel Rajaonarivonivelomanantsoa, Ruan de Kock +12

Reinforcement learning (RL) systems have countless applications, from energy-grid management to protein design. However, such real-world scenarios are often extremely difficult, co…

cs.AI2024★ 1 cited

Memory-Enhanced Neural Solvers for Routing Problems

Felix Chalumeau, Refiloe Shabe, Noah De Nicola +3

Routing Problems are central to many real-world applications, yet remain challenging due to their (NP-)hard nature. Amongst existing approaches, heuristics often offer the best tra…

cs.NE2024

Synergizing Quality-Diversity with Descriptor-Conditioned Reinforcement Learning

Maxence Faldor, Félix Chalumeau, Manon Flageat +1

A hallmark of intelligence is the ability to exhibit a wide range of effective behaviors. Inspired by this principle, Quality-Diversity algorithms, such as MAP-Elites, are evolutio…