1 paper
Jiacheng Wu, Yang Zhu, Hongye Su
For continuous-time linear quadratic regulation with unknown system matrices, data-driven off-policy policy iteration typically estimates the value matrix and the improved feedback…