2 papers
cs.AI2020
Suphx: Mastering Mahjong with Deep Reinforcement Learning
Junjie Li, Sotetsu Koyamada, Qiwei Ye +7
Artificial Intelligence (AI) has achieved great success in many domains, and game AI is widely regarded as its beachhead since the dawn of AI. In recent years, studies on game AI h…
stat.ML2017
Neural Sequence Model Training via -divergence Minimization
Sotetsu Koyamada, Yuta Kikuchi, Atsunori Kanemura +2
We propose a new neural sequence model training method in which the objective function is defined by -divergence. We demonstrate that the objective function generalizes the maxi…