4.3k citations · 4.8k across the 10 of their papers we have counts for
Showing 2020Show all
2 papers · 1 filter
cs.LG2020★ 150 cited
Scaling Laws for Autoregressive Generative Modeling
Tom Henighan, Jared Kaplan, Mor Katz +16
We identify empirical scaling laws for the cross-entropy loss in four domains: generative image modeling, video modeling, multimodal imagetext models, and mathemat…
cs.LG2020★ 49 cited
Phasic Policy Gradient
Karl Cobbe, Jacob Hilton, Oleg Klimov +1
We introduce Phasic Policy Gradient (PPG), a reinforcement learning framework which modifies traditional on-policy actor-critic methods by separating policy and value function trai…