692 citations · 3.4k across the 87 of their papers we have counts for
Showing 2017 · cs.LGShow all
2 papers · 2 filters
cs.LG2017★ 12 cited
Gradient-free Policy Architecture Search and Adaptation
Sayna Ebrahimi, Anna Rohrbach, Trevor Darrell
We develop a method for policy architecture search and adaptation via gradient-free optimization which can learn to perform autonomous driving tasks. By learning from both demonstr…
cs.LG2017
Curiosity-driven Exploration by Self-supervised Prediction
Deepak Pathak, Pulkit Agrawal, Alexei A. Efros +1
In many real-world scenarios, rewards extrinsic to the agent are extremely sparse, or absent altogether. In such cases, curiosity can serve as an intrinsic reward signal to enable…