19 citations · 22 across the 9 of their papers we have counts for
Showing 2019Show all
2 papers · 1 filter
cs.LG2019★ 1 cited
Optimistic Proximal Policy Optimization
Takahisa Imagawa, Takuya Hiraoka, Yoshimasa Tsuruoka
Reinforcement Learning, a machine learning framework for training an autonomous agent based on rewards, has shown outstanding results in various domains. However, it is known that…
cs.LG2019
Learning Robust Options by Conditional Value at Risk Optimization
Takuya Hiraoka, Takahisa Imagawa, Tatsuya Mori +2
Options are generally learned by using an inaccurate environment model (or simulator), which contains uncertain model parameters. While there are several methods to learn options t…