14 citations · 23 across the 12 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2024
When Do Off-Policy and On-Policy Policy Gradient Methods Align?
Davide Mambelli, Stephan Bongers, Onno Zoeter +2
Policy gradient methods are widely adopted reinforcement learning algorithms for tasks with continuous action spaces. These methods succeeded in many application domains, however,…
stat.ML2023
Fair Grading Algorithms for Randomized Exams
Jiale Chen, Jason Hartline, Onno Zoeter
This paper studies grading algorithms for randomized exams. In a randomized exam, each student is asked a small number of random questions from a large question bank. The predomina…