2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.AI2024
Game On: Towards Language Models as RL Experimenters
Jingwei Zhang, Thomas Lampe, Abbas Abdolmaleki +2
We propose an agent architecture that automates parts of the common reinforcement learning experiment workflow, to enable automated mastery of control domains for embodied agents.…
cs.LG2024★ 2 cited
Offline Actor-Critic Reinforcement Learning Scales to Large Models
Jost Tobias Springenberg, Abbas Abdolmaleki, Jingwei Zhang +9
We show that offline actor-critic reinforcement learning can scale to large models - such as transformers - and follows similar scaling laws as supervised learning. We find that of…