Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Boosting Soft Q-Learning by Bounding
Jacob Adamczyk, Volodymyr Makarenko, Stas Tiomkin +1
An agent's ability to leverage past experience is critical for efficiently solving new tasks. Prior work has focused on using value function estimates to obtain zero-shot approxima…
cs.LG2023
Bounding the Optimal Value Function in Compositional Reinforcement Learning
Jacob Adamczyk, Volodymyr Makarenko, Argenis Arriojas +2
In the field of reinforcement learning (RL), agents are often tasked with solving a variety of problems differing only in their reward functions. In order to quickly obtain solutio…