3 papers
cs.LG2026
Regret-Guided Search Control for Efficient Learning in AlphaZero
Yun-Jui Tsai, Wei-Yu Chen, Yan-Ru Ju +2
Reinforcement learning (RL) agents achieve remarkable performance but remain far less learning-efficient than humans. While RL agents require extensive self-play games to extract u…
cs.AI2024
Demystifying MuZero Planning: Interpreting the Learned Model
Hung Guei, Yan-Ru Ju, Wei-Yu Chen +1
MuZero has achieved superhuman performance in various games by using a dynamics network to predict the environment dynamics for planning, without relying on simulators. However, th…
cs.LG2024
Bridging Local and Global Knowledge via Transformer in Board Games
Yan-Ru Ju, Tai-Lin Wu, Chung-Chin Shih +1
Although AlphaZero has achieved superhuman performance in board games, recent studies reveal its limitations in handling scenarios requiring a comprehensive understanding of the en…