activity
20172022
most citedModel-Based Episodic Memory Induces Dynamic Hybrid Controls

3 citations · 9 across the 9 of their papers we have counts for

collaborators

12 papers

cs.LG2022

Uncertainty Aware System Identification with Universal Policies

Buddhika Laknath Semage, Thommen George Karimpanal, Santu Rana +1

Sim2real transfer is primarily concerned with transferring policies trained in simulation to potentially noisy real world environments. A common problem associated with sim2real tr…

cs.LG2022

Fast Model-based Policy Search for Universal Policy Networks

Buddhika Laknath Semage, Thommen George Karimpanal, Santu Rana +1

Adapting an agent's behaviour to new environments has been one of the primary focus areas of physics based reinforcement learning. Although recent approaches such as universal poli…

cs.LG20213 cited

Model-Based Episodic Memory Induces Dynamic Hybrid Controls

Hung Le, Thommen Karimpanal George, Majid Abdolshah +2

Episodic control enables sample efficiency in reinforcement learning by recalling past experiences from an episodic memory. We propose a new model-based episodic memory of trajecto…

cs.LG20211 cited

Balanced Q-learning: Combining the Influence of Optimistic and Pessimistic Targets

Thommen George Karimpanal, Hung Le, Majid Abdolshah +4

The optimistic nature of the Q-learning target leads to an overestimation bias, which is an inherent problem associated with standard learning. Such a bias fails to account for…

cs.LG2021

Plug and Play, Model-Based Reinforcement Learning

Majid Abdolshah, Hung Le, Thommen Karimpanal George +3

Sample-efficient generalisation of reinforcement learning approaches have always been a challenge, especially, for complex scenes with many components. In this work, we introduce P…

cs.LG20213 cited

A New Representation of Successor Features for Transfer across Dissimilar Environments

Majid Abdolshah, Hung Le, Thommen Karimpanal George +3

Transfer in reinforcement learning is usually achieved through generalisation across tasks. Whilst many studies have investigated transferring knowledge when the reward function ch…