Bayesian Additive Regression Trees using Bayesian Model Averaging.

Belinda Hernández,Stephen R Pennington,Andrew C Parnell,Adrian E Raftery

doi:10.1007/s11222-017-9767-1

Abstract

Bayesian Additive Regression Trees (BART) is a statistical sum of trees model. It can be considered a Bayesian version of machine learning tree ensemble methods where the individual trees are the base learners. However for datasets where the number of variables p is large the algorithm can become inefficient and computationally expensive. Another method which is popular for high dimensional data is random forests, a machine learning algorithm which grows trees using a greedy search for the best split points. However its default implementation does not produce probabilistic estimates or predictions. We propose an alternative fitting algorithm for BART called BART-BMA, which uses Bayesian Model Averaging and a greedy search algorithm to obtain a posterior distribution more efficiently than BART for datasets with large p. BART-BMA incorporates elements of both BART and random forests to offer a model-based algorithm which can deal with high-dimensional data. We have found that BART-BMA can be run in a reasonable time on a standard laptop for the "small n large p" scenario which is common in many areas of bioinformatics. We showcase this method using simulated data and data from two real proteomic experiments, one to distinguish between patients with cardiovascular disease and controls and another to classify aggressive from non-aggressive prostate cancer. We compare our results to their main competitors. Open source code written in R and Rcpp to run BART-BMA can be found at: https://github.com/BelindaHernandez/BART-BMA.git.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Bayesian Additive Regression Trees using Bayesian Model Averaging.

Abstract

Talk to us

Similar Papers

More From: Statistics and Computing

Lead the way for us

Journal: Statistics and Computing	Publication Date: Jul 27, 2017
Citations: 66

Similar Papers

Detection of Left Ventricular Hypertrophy Using Bayesian Additive Regression Trees: The MESA (Multi‐Ethnic Study of Atherosclerosis)
-
Journal of the American Heart Association | VOL. 8
--
04 May 2019
Journal of the American Heart Association | VOL. 8

Estimation of causal effects of multiple treatments in observational studies with a binary outcome.
Liangyuan Hu ... Michael Lopez
Statistical Methods in Medical Research | VOL. 29
Liangyuan Hu, et. al.Liangyuan Hu ... Michael Lopez
25 May 2020
Statistical Methods in Medical Research | VOL. 29

Genome-wide prediction using Bayesian additive regression trees
Patrik Waldmann
Genetics Selection Evolution | VOL. 48
Patrik WaldmannPatrik Waldmann
10 Jun 2016
Genetics Selection Evolution | VOL. 48

Model Mixing Using Bayesian Additive Regression Trees
John C. Yannotty ... Matthew T. Pratola
Technometrics | VOL. 66
John C. Yannotty, et. al.John C. Yannotty ... Matthew T. Pratola
16 Oct 2023
Technometrics | VOL. 66

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Bayesian Additive Regression Trees using Bayesian Model Averaging.

Abstract

Talk to us

Similar Papers

More From: Statistics and Computing