2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.AI2022★ 2 cited
Are AlphaZero-like Agents Robust to Adversarial Perturbations?
Li-Cheng Lan, Huan Zhang, Ti-Rong Wu +3
The success of AlphaZero (AZ) has demonstrated that neural-network-based Go AIs can surpass human performance by a large margin. Given that the state space of Go is extremely large…
cs.RO2022
Image-Based Conditioning for Action Policy Smoothness in Autonomous Miniature Car Racing with Reinforcement Learning
Bo-Jiun Hsu, Hoang-Giang Cao, I Lee +3
In recent years, deep reinforcement learning has achieved significant results in low-level controlling tasks. However, the problem of control smoothness has less attention. In auto…
cs.CV2018
Stochastic Gradient Descent with Hyperbolic-Tangent Decay on Classification
Bo Yang Hsueh, Wei Li, I-Chen Wu
Learning rate scheduler has been a critical issue in the deep neural network training. Several schedulers and methods have been proposed, including step decay scheduler, adaptive m…