48 citations · 48 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 48 cited
Scaling Up Models and Data with and
Adam Roberts, Hyung Won Chung, Anselm Levskaya +40
Recent neural network-based language models have benefited greatly from scaling up the size of training datasets and the number of parameters in the models themselves. Scaling can…
cs.LG2021
Q-Value Weighted Regression: Reinforcement Learning with Limited Data
Piotr Kozakowski, Łukasz Kaiser, Henryk Michalewski +2
Sample efficiency and performance in the offline setting have emerged as significant challenges of deep reinforcement learning. We introduce Q-Value Weighted Regression (QWR), a si…