3 papers
cs.LG2020
Model-Free Algorithm and Regret Analysis for MDPs with Long-Term Constraints
Qinbo Bai, Vaneet Aggarwal, Ather Gattami
In the optimization of dynamical systems, the variables typically have constraints. Such problems can be modeled as a constrained Markov Decision Process (CMDP). This paper conside…
math.OC2019
Escaping Saddle Points for Zeroth-order Nonconvex Optimization using Estimated Gradient Descent
Qinbo Bai, Mridul Agarwal, Vaneet Aggarwal
Gradient descent and its variants are widely used in machine learning. However, oracle access of gradient may not be available in many applications, limiting the direct use of grad…
eess.SP2019
Deep Learning based Channel Estimation Algorithm over Time Selective Fading Channels
Qinbo Bai, Jintao Wang, Yue Zhang +1
The research about deep learning application for physical layer has been received much attention in recent years. In this paper, we propose a Deep Learning (DL) based channel estim…