2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.LG2023
Handling Cost and Constraints with Off-Policy Deep Reinforcement Learning
Jared Markowitz, Jesse Silverberg, Gary Collins
By reusing data throughout training, off-policy deep reinforcement learning algorithms offer improved sample efficiency relative to on-policy approaches. For continuous action spac…
cs.LG2023★ 2 cited
Clipped-Objective Policy Gradients for Pessimistic Policy Optimization
Jared Markowitz, Edward W. Staley
To facilitate efficient learning, policy gradient approaches to deep reinforcement learning (RL) are typically paired with variance reduction measures and strategies for making lar…