3 citations · 4 across the 3 of their papers we have counts for
3 papers
Model-Based Episodic Memory Induces Dynamic Hybrid Controls
Hung Le, Thommen Karimpanal George, Majid Abdolshah +2
Episodic control enables sample efficiency in reinforcement learning by recalling past experiences from an episodic memory. We propose a new model-based episodic memory of trajecto…
Balanced Q-learning: Combining the Influence of Optimistic and Pessimistic Targets
Thommen George Karimpanal, Hung Le, Majid Abdolshah +4
The optimistic nature of the Q-learning target leads to an overestimation bias, which is an inherent problem associated with standard learning. Such a bias fails to account for…
Plug and Play, Model-Based Reinforcement Learning
Majid Abdolshah, Hung Le, Thommen Karimpanal George +3
Sample-efficient generalisation of reinforcement learning approaches have always been a challenge, especially, for complex scenes with many components. In this work, we introduce P…