6 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.LG2021
Q-Value Weighted Regression: Reinforcement Learning with Limited Data
Piotr Kozakowski, Łukasz Kaiser, Henryk Michalewski +2
Sample efficiency and performance in the offline setting have emerged as significant challenges of deep reinforcement learning. We introduce Q-Value Weighted Regression (QWR), a si…
cs.LG2019★ 6 cited
Uncertainty-sensitive Learning and Planning with Ensembles
Piotr Miłoś, Łukasz Kuciński, Konrad Czechowski +2
We propose a reinforcement learning framework for discrete environments in which an agent makes both strategic and tactical decisions. The former manifests itself through the use o…