6 citations · 21 across the 8 of their papers we have counts for
3 papers · 1 filter
Learning with Options that Terminate Off-Policy
Anna Harutyunyan, Peter Vrancx, Pierre-Luc Bacon +2
A temporally abstract action, or an option, is specified by a policy and a termination condition: the policy guides option behavior, and the termination condition roughly determine…
Reinforcement Learning in POMDPs with Memoryless Options and Option-Observation Initiation Sets
Denis Steckelmacher, Diederik M. Roijers, Anna Harutyunyan +3
Many real-world reinforcement learning problems have a hierarchical nature, and often exhibit some degree of partial observability. While hierarchy and partial observability are us…
Analysing Congestion Problems in Multi-agent Reinforcement Learning
Roxana Rădulescu, Peter Vrancx, Ann Nowé
Congestion problems are omnipresent in today's complex networks and represent a challenge in many research domains. In the context of Multi-agent Reinforcement Learning (MARL), app…