87 citations · 305 across the 29 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020
Maximum-and-Concatenation Networks
Xingyu Xie, Hao Kong, Jianlong Wu +3
While successful in many fields, deep neural networks (DNNs) still suffer from some open problems such as bad local minima and unsatisfactory generalization performance. In this wo…
cs.LG2018★ 2 cited
NADPEx: An on-policy temporally consistent exploration method for deep reinforcement learning
Sirui Xie, Junning Huang, Lanxin Lei +4
Reinforcement learning agents need exploratory behaviors to escape from local optima. These behaviors may include both immediate dithering perturbation and temporally consistent ex…