3 papers
cs.LG2025
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
Vinay Kanakeri, Shivam Bajaj, Ashwin Verma +2
It is known that reinforcement learning (RL) is data-hungry. To improve sample-efficiency of RL, it has been proposed that the learning algorithm utilize data from 'approximately s…
eess.SY2025
Outlier-Robust Linear System Identification Under Heavy-tailed Noise
Vinay Kanakeri, Aritra Mitra
We consider the problem of estimating the state transition matrix of a linear time-invariant (LTI) system, given access to multiple independent trajectories sampled from the system…
eess.SY2025
Boosting-Enabled Robust System Identification of Partially Observed LTI Systems Under Heavy-Tailed Noise
Vinay Kanakeri, Aritra Mitra
We consider the problem of system identification of partially observed linear time-invariant (LTI) systems. Given input-output data, we provide non-asymptotic guarantees for identi…