Optimization Landscape of Gradient Descent for Discrete-time Static Output Feedback
arXiv:2109.13132 · doi:10.23919/ACC53348.2022.9867384
Abstract
In this paper, we analyze the optimization landscape of gradient descent methods for static output feedback (SOF) control of discrete-time linear time-invariant systems with quadratic cost. The SOF setting can be quite common, for example, when there are unmodeled hidden states in the underlying process. We first establish several important properties of the SOF cost function, including coercivity, L-smoothness, and M-Lipschitz continuous Hessian. We then utilize these properties to show that the gradient descent is able to converge to a stationary point at a dimension-free rate. Furthermore, we prove that under some mild conditions, gradient descent converges linearly to a local minimum if the starting point is close to one. These results not only characterize the performance of gradient descent in optimizing the SOF problem, but also shed light on the efficiency of general policy gradient methods in reinforcement learning.
References in corpus (6)
- Distributional Soft Actor-Critic: Off-Policy Reinforcement Learning for Addressing Value Estimation Errors
- How to Escape Saddle Points Efficiently
- Hierarchical Reinforcement Learning for Self-Driving Decision-Making without Reliance on Labeled Driving Data
- Analysis of the Optimization Landscape of Linear Quadratic Gaussian (LQG) Control
- Optimization Landscape of Gradient Descent for Discrete-time Static Output Feedback
- Policy Optimization for Markovian Jump Linear Quadratic Control: Gradient-Based Methods and Global Convergence
Cited by in corpus (4)
- On the Optimization Landscape of Dynamic Output Feedback Linear Quadratic Control
- Optimization Landscape of Gradient Descent for Discrete-time Static Output Feedback
- On the Optimization Landscape of Dynamic Output Feedback: A Case Study for Linear Quadratic Regulator
- Lossless Convexification and Duality