3 papers
cs.AI2023
CFR-p: Counterfactual Regret Minimization with Hierarchical Policy Abstraction, and its Application to Two-player Mahjong
Shiheng Wang
Counterfactual Regret Minimization(CFR) has shown its success in Texas Hold'em poker. We apply this algorithm to another popular incomplete information game, Mahjong. Compared to t…
cs.GT2019
Pure Strategy Best Responses to Mixed Strategies in Repeated Games
Shiheng Wang, Fangzhen Lin
Repeated games are difficult to analyze, especially when agents play mixed strategies. We study one-memory strategies in iterated prisoner's dilemma, then generalize the result to…
cs.GT2017
Invincible Strategies of Iterated Prisoner's Dilemma
Shiheng Wang, Fangzhen Lin
Iterated Prisoner's Dilemma(IPD) is a well-known benchmark for studying the long term behaviors of rational agents, such as how cooperation can emerge among selfish and unrelated a…