1 citations · 1 across the 4 of their papers we have counts for
4 papers
Energy Management in a Cooperative Energy Harvesting Wireless Sensor Network
Arghyadeep Barat, Prabuchandran. K. J, Shalabh Bhatnagar
In this paper, we consider the problem of finding an optimal energy management policy for a network of sensor nodes capable of harvesting their own energy and sharing it with other…
The Reinforce Policy Gradient Algorithm Revisited
Shalabh Bhatnagar
We revisit the Reinforce policy gradient algorithm from the literature. Note that this algorithm typically works with cost returns obtained over random length episodes obtained fro…
A Framework for Provably Stable and Consistent Training of Deep Feedforward Networks
Arunselvan Ramaswamy, Shalabh Bhatnagar, Naman Saxena
We present a novel algorithm for training deep neural networks in supervised (classification and regression) and unsupervised (reinforcement learning) scenarios. This algorithm com…
A Cubic-regularized Policy Newton Algorithm for Reinforcement Learning
Mizhaan Prajit Maniyar, Akash Mondal, Prashanth L. A. +1
We consider the problem of control in the setting of reinforcement learning (RL), where model information is not available. Policy gradient algorithms are a popular solution approa…