21 citations · 21 across the 4 of their papers we have counts for
3 papers · 1 filter
Offline-Online Reinforcement Learning for Linear Mixture MDPs
Zhongjun Zhang, Sean R. Sinclair
We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data are collected by an unknown be…
Adaptive Discretization for Model-Based Reinforcement Learning
Sean R. Sinclair, Tianyu Wang, Gauri Jain +2
We introduce the technique of adaptive discretization to design an efficient model-based episodic reinforcement learning algorithm in large (potentially continuous) state-action sp…
Adaptive Discretization for Episodic Reinforcement Learning in Metric Spaces
Sean R. Sinclair, Siddhartha Banerjee, Christina Lee Yu
We present an efficient algorithm for model-free episodic reinforcement learning on large (potentially continuous) state-action spaces. Our algorithm is based on a novel -learni…