Showing cs.LGShow all
2 papers · 1 filter
cs.LG2021
Offline Policy Comparison under Limited Historical Agent-Environment Interactions
Anton Dereventsov, Joseph D. Daws, Clayton Webster
We address the challenge of policy evaluation in real-world applications of reinforcement learning systems where the available historical data is limited due to ethical, practical,…
cs.LG2019
Neural network integral representations with the ReLU activation function
Armenak Petrosyan, Anton Dereventsov, Clayton Webster
In this effort, we derive a formula for the integral representation of a shallow neural network with the ReLU activation function. We assume that the outer weighs admit a finite $L…