2 citations · 2 across the 5 of their papers we have counts for
4 papers · 1 filter
In-Context Analogical Reasoning with Pre-Trained Language Models
Xiaoyang Hu, Shane Storks, Richard L. Lewis +1
Analogical reasoning is a fundamental capacity of human cognition that allows us to reason abstractly about novel situations by relating them to past experiences. While it is thoug…
How Should an Agent Practice?
Janarthanan Rajendran, Richard Lewis, Vivek Veeriah +2
We present a method for learning intrinsic reward functions to drive the learning of an agent during periods of practice in which extrinsic task rewards are not available. During p…
Discovery of Useful Questions as Auxiliary Tasks
Vivek Veeriah, Matteo Hessel, Zhongwen Xu +6
Arguably, intelligent agents ought to be able to discover their own questions so that in learning answers for them they learn unanticipated useful knowledge and skills; this depart…
Deep Learning for Reward Design to Improve Monte Carlo Tree Search in ATARI Games
Xiaoxiao Guo, Satinder Singh, Richard Lewis +1
Monte Carlo Tree Search (MCTS) methods have proven powerful in planning for sequential decision-making problems such as Go and video games, but their performance can be poor when t…