2 papers
cs.LG2026
When Do Options Help? Policy Necrosis and Redundant Coverage in Option-Critic
Bingyun Liu, Yuheng Jing
Option-critic learns options: sub-policies together with a learned rule for when each one hands control back. Its headline result is that performance improves as options are added.…
cs.LG2024
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
Hang Xu, Kai Li, Bingyun Liu +4
Counterfactual regret minimization (CFR) is a family of algorithms for effectively solving imperfect-information games. It decomposes the total regret into counterfactual regrets,…