POMCP with Human Preferences in Settlers of Catan

Mihai Dobre,Alex Lascarides

doi:10.1609/aiide.v14i1.13014

POMCP with Human Preferences in Settlers of Catan

Mihai Dobre, Alex Lascarides

Open Access

https://doi.org/10.1609/aiide.v14i1.13014

Copy DOI

Journal: Proceedings of the AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment	Publication Date: Sep 25, 2018
Citations: 1

Affiliation: University of Edinburgh

#High-level Strategies #Human Preferences + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

We present a suite of techniques for extending the Partially Observable Monte Carlo Planning algorithm to handle complex multi-agent games. We design the planning algorithm to exploit the inherent structure of the game. When game rules naturally cluster the actions into sets called types, these can be leveraged to extract characteristics and high-level strategies from a sparse corpus of human play. Another key insight is to account for action legality both when extracting policies from game play and when these are used to inform the forward sampling method. We evaluate our algorithm against other baselines and versus ablated versions of itself in the well-known board game Settlers of Catan.

Full Text