7 citations · 7 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
Composing Efficient, Robust Tests for Policy Selection
Dustin Morrill, Thomas J. Walsh, Daniel Hernandez +2
Modern reinforcement learning systems produce many high-quality policies throughout the learning process. However, to choose which policy to actually deploy in the real world, they…
cs.LG2012★ 7 cited
Exploring compact reinforcement-learning representations with linear regression
Thomas J. Walsh, Istvan Szita, Carlos Diuk +1
This paper presents a new algorithm for online linear regression whose efficiency guarantees satisfy the requirements of the KWIK (Knows What It Knows) framework. The algorithm imp…