1 citations · 1 across the 2 of their papers we have counts for
3 papers
QT-TDM: Planning With Transformer Dynamics Model and Autoregressive Q-Learning
Mostafa Kotb, Cornelius Weber, Muhammad Burhan Hafez +1
Inspired by the success of the Transformer architecture in natural language processing and computer vision, we investigate the use of Transformers in Reinforcement Learning (RL), s…
Model Predictive Control with Self-supervised Representation Learning
Jonas Matthies, Muhammad Burhan Hafez, Mostafa Kotb +1
Over the last few years, we have not seen any major developments in model-free or model-based learning methods that would make one obsolete relative to the other. In most cases, th…
Sample-efficient Real-time Planning with Curiosity Cross-Entropy Method and Contrastive Learning
Mostafa Kotb, Cornelius Weber, Stefan Wermter
Model-based reinforcement learning (MBRL) with real-time planning has shown great potential in locomotion and manipulation control tasks. However, the existing planning methods, su…