Value functions for depth-limited solving in zero-sum imperfect-information games

Vojtěch Kovařík,Dominik Seitz,Viliam Lisý,Jan Rudolf,Shuo Sun,Karel Ha

doi:10.1016/j.artint.2022.103805

Vojtěch Kovařík, Dominik Seitz + Show 4 more

Open Access

https://doi.org/10.1016/j.artint.2022.103805

Copy DOI

Abstract

We provide a formal definition of depth-limited games together with an accessible and rigorous explanation of the underlying concepts, both of which were previously missing in imperfect-information games. The definition works for an arbitrary (perfect recall) extensive-form game and is not tied to any specific game-solving algorithm. Moreover, this framework unifies and significantly extends three approaches to depth-limited solving that previously existed in extensive-form games and multiagent reinforcement learning but were not known to be compatible. A key ingredient of these depth-limited games is value functions. Focusing on two-player zero-sum imperfect-information games, we show how to obtain optimal value functions and prove that public information provides both necessary and sufficient context for computing them. We provide a domain-independent encoding of the domains that allows for approximating value functions even by simple feed-forward neural networks, which are then able to generalize to unseen parts of the game. We use the resulting value network to implement a depth-limited version of counterfactual regret minimization. In three distinct domains, we show that the algorithm's exploitability is roughly linearly dependent on the value network's quality and that it is not difficult to train a value network with which depth-limited CFR's performance is as good as that of CFR with access to the full game.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Value functions for depth-limited solving in zero-sum imperfect-information games

Abstract

Talk to us

Similar Papers

More From: Artificial Intelligence

Lead the way for us

Journal: Artificial Intelligence	Publication Date: Oct 19, 2022
Citations: 4

Similar Papers

Finding Equilibria in Games of No Chance
Kristoffer Arnsfelt Hansen ... Peter Bro Miltersen
-
Kristoffer Arnsfelt Hansen, et. al.Kristoffer Arnsfelt Hansen ... Peter Bro Miltersen
16 Jul 2007
16 Jul 2007

Rethinking Formal Models of Partially Observable Multiagent Decision Making (Extended Abstract)
Vojtěch Kovařík ... Michael Bowling
-
Vojtěch Kovařík, et. al.Vojtěch Kovařík ... Michael Bowling
01 Aug 2023
01 Aug 2023

Rethinking formal models of partially observable multiagent decision making
Vojtěch Kovařík ... Viliam Lisý
Artificial Intelligence | VOL. 303
Vojtěch Kovařík, et. al.Vojtěch Kovařík ... Viliam Lisý
26 Nov 2021
Artificial Intelligence | VOL. 303

Double-oracle algorithm for computing an exact nash equilibrium in zero-sum extensive-form games
...
-
, et. al. ...
27 Jun 2013
27 Jun 2013

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Value functions for depth-limited solving in zero-sum imperfect-information games

Abstract

Talk to us

Similar Papers

More From: Artificial Intelligence