4 papers · 1 filter
Offline Action-Free Learning of Ex-BMDPs by Comparing Diverse Datasets
Alexander Levine, Peter Stone, Amy Zhang
While sequential decision-making environments often involve high-dimensional observations, not all features of these observations are relevant for control. In particular, the obser…
Learning a Fast Mixing Exogenous Block MDP using a Single Trajectory
Alexander Levine, Peter Stone, Amy Zhang
In order to train agents that can quickly adapt to new objectives or reward functions, efficient unsupervised representation learning in sequential decision-making environments can…
Proto Successor Measure: Representing the Behavior Space of an RL Agent
Siddhant Agarwal, Harshit Sikchi, Peter Stone +1
Having explored an environment, intelligent agents should be able to transfer their knowledge to most downstream tasks within that environment without additional interactions. Refe…
Multistep Inverse Is Not All You Need
Alexander Levine, Peter Stone, Amy Zhang
In real-world control settings, the observation space is often unnecessarily high-dimensional and subject to time-correlated noise. However, the controllable dynamics of the system…