3 papers
cs.AI2024
Boltzmann State-Dependent Rationality
Osher Lerner
This paper expands on existing learned models of human behavior via a measured step in structured irrationality. Specifically, by replacing the suboptimality constant in a Bolt…
cs.RO2024
Precise Object Placement Using Force-Torque Feedback
Osher Lerner, Zachary Tam, Michael Equi
Precise object manipulation and placement is a common problem for household robots, surgery robots, and robots working on in-situ construction. Prior work using computer vision, de…
cs.LG2018
Natural Gradient Deep Q-learning
Ethan Knight, Osher Lerner
We present a novel algorithm to train a deep Q-learning agent using natural-gradient techniques. We compare the original deep Q-network (DQN) algorithm to its natural-gradient coun…