Automatic Induction of Bellman-Error Features for Probabilistic Planning

J Wu,R Givan

doi:10.1613/jair.3021

Abstract

Domain-specific features are important in representing problem structure throughout machine learning and decision-theoretic planning. In planning, once state features are provided, domain-independent algorithms such as approximate value iteration can learn weighted combinations of those features that often perform well as heuristic estimates of state value (e.g., distance to the goal). Successful applications in real-world domains often require features crafted by human experts. Here, we propose automatic processes for learning useful domain-specific feature sets with little or no human intervention. Our methods select and add features that describe state-space regions of high inconsistency in the Bellman equation (statewise Bellman error) during approximate value iteration. Our method can be applied using any real-valued-feature hypothesis space and corresponding learning method for selecting features from training sets of state-value pairs. We evaluate the method with hypothesis spaces defined by both relational and propositional feature languages, using nine probabilistic planning domains. We show that approximate value iteration using a relational feature space performs at the state-of-the-art in domain-independent stochastic relational planning. Our method provides the first domain-independent approach that plays Tetris successfully (without human-engineered features).

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Journal of Artificial Intelligence Research	Publication Date: Aug 30, 2010
Citations: 43	License type: cc-by

R Discovery Prime

R Discovery Prime

Automatic Induction of Bellman-Error Features for Probabilistic Planning

Abstract

Talk to us

Similar Papers

More From: Journal of Artificial Intelligence Research

Lead the way for us

Similar Papers

Feature-Discovering Approximate Value Iteration Methods
Jia-Hong Wu ... Robert Givan
-
Jia-Hong Wu, et. al.Jia-Hong Wu ... Robert Givan
01 Jan 2004
01 Jan 2004

Efficient incremental planning and learning with multi-valued decision diagrams
Jean-Christophe Magnan ... Pierre-Henri Wuillemin
Journal of Applied Logic | VOL. 22
Jean-Christophe Magnan, et. al.Jean-Christophe Magnan ... Pierre-Henri Wuillemin
17 Nov 2016
Journal of Applied Logic | VOL. 22

Secure Linear Quadratic Regulator Using Sparse Model-Free Reinforcement Learning
Bahare Kiumarsi ... Tamer Basar
-
Bahare Kiumarsi, et. al.Bahare Kiumarsi ... Tamer Basar
01 Dec 2019
01 Dec 2019

TH‐C‐137‐11: Dose‐Guided Automatic IMRT Planning: A Feasibility Study
Y Sheng ... T Li
Medical Physics | VOL. 40
Y Sheng, et. al.Y Sheng ... T Li
01 Jun 2013
TH‐C‐137‐11: Dose‐Guided Automatic IMRT Planning: A Feasibility Study
Y Sheng ... T Li

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Automatic Induction of Bellman-Error Features for Probabilistic Planning

Abstract

Talk to us

Similar Papers

More From: Journal of Artificial Intelligence Research