1 paper · 1 filter
Leonardo F. Toso, Han Wang, James Anderson
We address the problem of designing an LQR controller in a distributed setting, where M similar but not identical systems share their locally computed policy gradient (PG) estimate…