3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.LG2020★ 3 cited
The act of remembering: a study in partially observable reinforcement learning
Rodrigo Toro Icarte, Richard Valenzano, Toryn Q. Klassen +3
Reinforcement Learning (RL) agents typically learn memoryless policies---policies that only consider the last observation when selecting actions. Learning memoryless policies is ef…
cs.AI2017★ 1 cited
A Formal Characterization of the Local Search Topology of the Gap Heuristic
Richard Anthony Valenzano, Danniel Sihui Yang
The pancake puzzle is a classic optimization problem that has become a standard benchmark for heuristic search algorithms. In this paper, we provide full proofs regarding the local…