7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 7 cited
Measuring Sample Efficiency and Generalization in Reinforcement Learning Benchmarks: NeurIPS 2020 Procgen Benchmark
Sharada Mohanty, Jyotish Poonganam, Adrien Gaidon +20
The NeurIPS 2020 Procgen Competition was designed as a centralized benchmark with clearly defined tasks for measuring Sample Efficiency and Generalization in Reinforcement Learning…
cs.LG2019
Multi-Preference Actor Critic
Ishan Durugkar, Matthew Hausknecht, Adith Swaminathan +1
Policy gradient algorithms typically combine discounted future rewards with an estimated value function, to compute the direction and magnitude of parameter updates. However, for m…