4k citations
- Nvidia (United States)US22 papers
- University of TorontoCA11 papers
- Stanford UniversityUS8 papers
- University of California, MercedUS6 papers
- University of WashingtonUS6 papers
- Massachusetts Institute of TechnologyUS5 papers
- Duke UniversityUS4 papers
- Google (United States)US4 papers
- Nanyang Technological UniversitySG4 papers
- Texas A&M UniversityUS4 papers
- University of CopenhagenDK4 papers
- Imperial College LondonGB3 papers
Showing 2016Show all
2 papers · 1 filter
cs.LG2016★ 28 cited
Reinforcement Learning through Asynchronous Advantage Actor-Critic on a GPU
Mohammad Babaeizadeh, Iuri Frosio, Stephen Tyree +2
We introduce a hybrid CPU/GPU version of the Asynchronous Advantage Actor-Critic (A3C) algorithm, currently the state-of-the-art method in reinforcement learning for various gaming…
cs.CV2016
SEBOOST - Boosting Stochastic Learning Using Subspace Optimization Techniques
Elad Richardson, Rom Herskovitz, Boris Ginsburg +1
We present SEBOOST, a technique for boosting the performance of existing stochastic optimization methods. SEBOOST applies a secondary optimization process in the subspace spanned b…