Rethinking Formal Models of Partially Observable Multiagent Decision Making

Kovařík, Vojtěch; Schmid, Martin; Burch, Neil; Bowling, Michael; Lisý, Viliam

Computer Science > Artificial Intelligence

arXiv:1906.11110 (cs)

[Submitted on 26 Jun 2019 (v1), last revised 28 Sep 2021 (this version, v4)]

Title:Rethinking Formal Models of Partially Observable Multiagent Decision Making

Authors:Vojtěch Kovařík, Martin Schmid, Neil Burch, Michael Bowling, Viliam Lisý

View PDF

Abstract:Multiagent decision-making in partially observable environments is usually modelled as either an extensive-form game (EFG) in game theory or a partially observable stochastic game (POSG) in multiagent reinforcement learning (MARL). One issue with the current situation is that while most practical problems can be modelled in both formalisms, the relationship of the two models is unclear, which hinders the transfer of ideas between the two communities. A second issue is that while EFGs have recently seen significant algorithmic progress, their classical formalization is unsuitable for efficient presentation of the underlying ideas, such as those around decomposition.
To solve the first issue, we introduce factored-observation stochastic games (FOSGs), a minor modification of the POSG formalism which distinguishes between private and public observation and thereby greatly simplifies decomposition. To remedy the second issue, we show that FOSGs and POSGs are naturally connected to EFGs: by "unrolling" a FOSG into its tree form, we obtain an EFG. Conversely, any perfect-recall timeable EFG corresponds to some underlying FOSG in this manner. Moreover, this relationship justifies several minor modifications to the classical EFG formalization that recently appeared as an implicit response to the model's issues with decomposition. Finally, we illustrate the transfer of ideas between EFGs and MARL by presenting three key EFG techniques -- counterfactual regret minimization, sequence form, and decomposition -- in the FOSG framework.

Comments:	A 2020 update of the original 2019 version of the paper. (Rewrote the main text and clarified the relationship between FOSGs/POSGs and EFGs. Some of the technical results are now presented in the appendix.)
Subjects:	Artificial Intelligence (cs.AI); Computer Science and Game Theory (cs.GT)
Cite as:	arXiv:1906.11110 [cs.AI]
	(or arXiv:1906.11110v4 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.1906.11110

Submission history

From: Vojtech Kovarik [view email]
[v1] Wed, 26 Jun 2019 14:01:34 UTC (502 KB)
[v2] Wed, 30 Sep 2020 10:50:22 UTC (1,238 KB)
[v3] Mon, 26 Oct 2020 12:20:13 UTC (1,238 KB)
[v4] Tue, 28 Sep 2021 16:30:55 UTC (1,299 KB)

Computer Science > Artificial Intelligence

Title:Rethinking Formal Models of Partially Observable Multiagent Decision Making

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Rethinking Formal Models of Partially Observable Multiagent Decision Making

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators