1.3k citations · 2.3k across the 66 of their papers we have counts for
4 papers · 1 filter
Experience Replay with Likelihood-free Importance Weights
Samarth Sinha, Jiaming Song, Animesh Garg +1
The use of past experiences to accelerate temporal difference (TD) learning of value functions, or experience replay, is a key component in deep reinforcement learning. Prioritizat…
Deterministic Policy Optimization by Combining Pathwise and Score Function Estimators for Discrete Action Spaces
Daniel Levy, Stefano Ermon
Policy optimization methods have shown great promise in solving complex reinforcement and imitation learning tasks. While model-free methods are broadly applicable, they often requ…
A Survey of Human Activity Recognition Using WiFi CSI
Siamak Yousefi, Hirokazu Narui, Sankalp Dayal +2
In this article, we present a survey of recent advances in passive human behaviour recognition in indoor areas using the channel state information (CSI) of commercial WiFi systems.…
Uniform Solution Sampling Using a Constraint Solver As an Oracle
Stefano Ermon, Carla P. Gomes, Bart Selman
We consider the problem of sampling from solutions defined by a set of hard constraints on a combinatorial space. We propose a new sampling technique that, while enforcing a unifor…