4 citations · 7 across the 3 of their papers we have counts for
1 paper · 1 filter
Jost Tobias Springenberg, Abbas Abdolmaleki, Jingwei Zhang +9
We show that offline actor-critic reinforcement learning can scale to large models - such as transformers - and follows similar scaling laws as supervised learning. We find that of…