6 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.LG2019★ 6 cited
A Tale of Two-Timescale Reinforcement Learning with the Tightest Finite-Time Bound
Gal Dalal, Balazs Szorenyi, Gugan Thoppe
Policy evaluation in reinforcement learning is often conducted using two-timescale stochastic approximation, which results in various gradient temporal difference methods such as G…
cs.LG2016★ 1 cited
Supervised Learning for Optimal Power Flow as a Real-Time Proxy
Raphael Canyasse, Gal Dalal, Shie Mannor
In this work we design and compare different supervised learning algorithms to compute the cost of Alternating Current Optimal Power Flow (ACOPF). The motivation for quick calculat…