activity
20172022
most citedA Meta Reinforcement Learning Approach for Predictive Autoscaling in the Cloud

48 citations · 90 across the 8 of their papers we have counts for

collaborators

14 papers

cs.LG202248 cited

A Meta Reinforcement Learning Approach for Predictive Autoscaling in the Cloud

Siqiao Xue, Chao Qu, Xiaoming Shi +11

Predictive autoscaling (autoscaling with workload forecasting) is an important mechanism that supports autonomous adjustment of computing resources in accordance with fluctuating w…

math.OC2022

Convex Parameterization of Stabilizing Controllers and its LMI-based Computation via Filtering

Mauricio C. de Oliveira, Yang Zheng

Various new implicit parameterizations for stabilizing controllers that allow one to impose structural constraints on the controller have been proposed lately. They are convex but…

math.OC202118 cited

Analysis of the Optimization Landscape of Linear Quadratic Gaussian (LQG) Control

Yang Zheng, Yujie Tang, Na Li

This paper revisits the classical Linear Quadratic Gaussian (LQG) control from a modern optimization perspective. We analyze two aspects of the optimization landscape of the LQG pr…

math.OC2020

Sample Complexity of Linear Quadratic Gaussian (LQG) Control for Output Feedback Systems

Yang Zheng, Luca Furieri, Maryam Kamgarpour +1

This paper studies a class of partially observed Linear Quadratic Gaussian (LQG) problems with unknown dynamics. We establish an end-to-end sample complexity bound on learning a ro…

eess.SY2020

Impact of Disturbances on Mixed Traffic Control with Autonomous Vehicles

Ross Drummond, Yang Zheng

This paper investigates the impact of disturbances on controlling an autonomous vehicle to smooth mixed traffic flow in a ring road setup. By exploiting the ring structure of this…

cs.LG20205 cited

BI-MAML: Balanced Incremental Approach for Meta Learning

Yang Zheng, Jinlin Xiang, Kun Su +1

We present a novel Balanced Incremental Model Agnostic Meta Learning system (BI-MAML) for learning multiple tasks. Our method implements a meta-update rule to incrementally adapt i…