11 citations · 27 across the 12 of their papers we have counts for
4 papers · 1 filter
AI Alignment: A Comprehensive Survey
Jiaming Ji, Tianyi Qiu, Boyuan Chen +23
AI alignment aims to make AI systems behave in line with human intentions and values. As AI systems grow more capable, so do risks from misalignment. To provide a comprehensive and…
JiangJun: Mastering Xiangqi by Tackling Non-Transitivity in Two-Player Zero-Sum Games
Yang Li, Kun Xiong, Yingping Zhang +6
This paper presents an empirical exploration of non-transitivity in perfect-information games, specifically focusing on Xiangqi, a traditional Chinese board game comparable in game…
Neural Auto-Curricula
Xidong Feng, Oliver Slumbers, Ziyu Wan +5
When solving two-player zero-sum games, multi-agent reinforcement learning (MARL) algorithms often create populations of agents where, at each iteration, a new agent is discovered…
Solving the Rubik's Cube Without Human Knowledge
Stephen McAleer, Forest Agostinelli, Alexander Shmakov +1
A generally intelligent agent must be able to teach itself how to solve problems in complex domains with minimal human supervision. Recently, deep reinforcement learning algorithms…