3 papers
cs.AI2026
Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX
Soichiro Nishimori, Shinri Okano, Keigo Habara +3
Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes present a unique combination of chal…
cs.GT2023
Convergence analysis and acceleration of the smoothing methods for solving extensive-form games
Keigo Habara, Ellen Hidemi Fukuda, Nobuo Yamashita
The extensive-form game has been studied considerably in recent years. It can represent games with multiple decision points and incomplete information, and hence it is helpful in f…
cs.AI2023
Pgx: Hardware-Accelerated Parallel Game Simulators for Reinforcement Learning
Sotetsu Koyamada, Shinri Okano, Soichiro Nishimori +4
We propose Pgx, a suite of board game reinforcement learning (RL) environments written in JAX and optimized for GPU/TPU accelerators. By leveraging JAX's auto-vectorization and par…