Bayesian Variable Selection in Clustering High-Dimensional Data

Mahlet G Tadesse,Naijun Sha,Marina Vannucci

doi:10.1198/016214504000001565

Abstract

Over the last decade, technological advances have generated an explosion of data with substantially smaller sample size relative to the number of covariates (p ≫ n). A common goal in the analysis of such data involves uncovering the group structure of the observations and identifying the discriminating variables. In this article we propose a methodology for addressing these problems simultaneously. Given a set of variables, we formulate the clustering problem in terms of a multivariate normal mixture model with an unknown number of components and use the reversible-jump Markov chain Monte Carlo technique to define a sampler that moves between different dimensional spaces. We handle the problem of selecting a few predictors among the prohibitively vast number of variable subsets by introducing a binary exclusion/inclusion latent vector, which gets updated via stochastic search techniques. We specify conjugate priors and exploit the conjugacy by integrating out some of the parameters. We describe strategies for posterior inference and explore the performance of the methodology with simulated and real datasets.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Bayesian Variable Selection in Clustering High-Dimensional Data

Abstract

Talk to us

Similar Papers

More From: Journal of the American Statistical Association

Lead the way for us

Journal: Journal of the American Statistical Association	Publication Date: Jun 1, 2005
Citations: 241

Similar Papers

Bayesian variable selection in clustering high-dimensional data via a mixture of finite mixtures
Woojin Doo ... Heeyoung Kim
Journal of Statistical Computation and Simulation | VOL. 91
Woojin Doo, et. al.Woojin Doo ... Heeyoung Kim
30 Mar 2021
Journal of Statistical Computation and Simulation | VOL. 91

Bayesian variable selection and computation for generalized linear models with conjugate priors
Ming-Hui Chen ... Joseph G Ibrahim
Bayesian Analysis | VOL. 3
Ming-Hui Chen, et. al.Ming-Hui Chen ... Joseph G Ibrahim
01 Sep 2008
Bayesian Analysis | VOL. 3

Perfect Simulation for Mixtures with Known and Unknown Number of Components
Sabyasachi Mukhopadhyay ... Sourabh Bhattacharya
Bayesian Analysis | VOL. 7
Sabyasachi Mukhopadhyay, et. al.Sabyasachi Mukhopadhyay ... Sourabh Bhattacharya
01 Sep 2012
Bayesian Analysis | VOL. 7

Estimation for flexible Weibull extension under progressive Type-II censoring
Sanjay Kumar Singh ... Umesh Singh
Journal of Data Science | VOL. 13
Sanjay Kumar Singh, et. al.Sanjay Kumar Singh ... Umesh Singh
07 Mar 2021
Journal of Data Science | VOL. 13

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Bayesian Variable Selection in Clustering High-Dimensional Data

Abstract

Talk to us

Similar Papers

More From: Journal of the American Statistical Association