1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Dylan Bates
Using the policy gradient algorithm, we train a single-hidden-layer neural network to balance a physically accurate simulation of a single inverted pendulum. The trained weights an…