2 papers
stat.ML2025
Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning
Jia Wan, Sean R. Sinclair, Devavrat Shah +1
We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endogenous components. Exogenous states evol…
math.OC2025
Multi-Objective LQR with Linear Scalarization
Ali Jadbabaie, Devavrat Shah, Sean R. Sinclair
We study finite approximation of the Pareto front for the multi-objective linear quadratic regulator (LQR). Classical results characterize Pareto-optimal controllers through weight…