Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:No-Regret Learning in Extensive-Form Games with Imperfect Recall

May 03, 2012

Marc Lanctot, Richard Gibson, Neil Burch, Martin Zinkevich, Michael Bowling

Figure 1 for No-Regret Learning in Extensive-Form Games with Imperfect Recall

Figure 2 for No-Regret Learning in Extensive-Form Games with Imperfect Recall

Figure 3 for No-Regret Learning in Extensive-Form Games with Imperfect Recall

Share this with someone who'll enjoy it:

Abstract:Counterfactual Regret Minimization (CFR) is an efficient no-regret learning algorithm for decision problems modeled as extensive games. CFR's regret bounds depend on the requirement of perfect recall: players always remember information that was revealed to them and the order in which it was revealed. In games without perfect recall, however, CFR's guarantees do not apply. In this paper, we present the first regret bound for CFR when applied to a general class of games with imperfect recall. In addition, we show that CFR applied to any abstraction belonging to our general class results in a regret bound not just for the abstract game, but for the full game as well. We verify our theory and show how imperfect recall can be used to trade a small increase in regret for a significant reduction in memory in three domains: die-roll poker, phantom tic-tac-toe, and Bluff.

* 21 pages, 4 figures, expanded version of article to appear in Proceedings of the Twenty-Ninth International Conference on Machine Learning

View paper on

Share this with someone who'll enjoy it:

Title:No-Regret Learning in Extensive-Form Games with Imperfect Recall

Paper and Code