New feature selection for gene expression classification based on degree of class overlap in principal dimensions

Somsak Rakkeitwinai,Chidchanok Lursinsap,Chatchawit Aporntewan,Apiwat Mutirangura

doi:10.1016/j.compbiomed.2015.01.022

Abstract

Micro-array data are typically characterized by high dimensional features with a small number of samples. Several problems in identifying genes causing diseases from micro-array data can be transformed into the problem of classifying the features extracted from gene expression in micro-array data. However, too many features can cause low prediction accuracy as well as high computational complexity. Dimensional reduction is a method to eliminate irrelevant features to improve the prediction accuracy. Typically, the eigenvalues or dimensional data variance from principal component analysis are used as criteria to select relevant features. This approach is simple but not efficient since it does not concern the degree of data overlap in each dimension in the feature space. A new method to select relevant features based on degree of dimensional data overlap with proper feature selection was introduced. Furthermore, our study concentrated on small sized data sets which usually occur in reality. The experimental results signified that this new approach can achieve substantially higher prediction accuracy when compared with other methods.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

New feature selection for gene expression classification based on degree of class overlap in principal dimensions

Abstract

Talk to us

Similar Papers

More From: Computers in Biology and Medicine

Lead the way for us

Journal: Computers in Biology and Medicine	Publication Date: Feb 7, 2015
Citations: 45

Similar Papers

Dimension reduction methods for microarray data: a review
Rabia Aziz ... Namita Srivastava
AIMS Bioengineering | VOL. 4
Rabia Aziz, et. al.Rabia Aziz ... Namita Srivastava
01 Jan 2017
AIMS Bioengineering | VOL. 4

Rasch-based high-dimensionality data reduction and class prediction with applications to microarray gene expression data
Andrej Kastrin ... Borut Peterlin
Expert Systems with Applications | VOL. 37
Andrej Kastrin, et. al.Andrej Kastrin ... Borut Peterlin
04 Jan 2010
Expert Systems with Applications | VOL. 37

Accurate detection of aneuploidies in array CGH and gene expression microarray data.
Chad L Myers ... Maitreya J Dunham
Bioinformatics | VOL. 20
Chad L Myers, et. al.Chad L Myers ... Maitreya J Dunham
29 Jul 2004
Bioinformatics | VOL. 20

Mean, median and tri-mean based statistical detection methods for differential gene expression in microarray data
Zhaohua Ji ... Chong Xing
-
Zhaohua Ji, et. al.Zhaohua Ji ... Chong Xing
01 Oct 2010
01 Oct 2010

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

New feature selection for gene expression classification based on degree of class overlap in principal dimensions

Abstract

Talk to us

Similar Papers

More From: Computers in Biology and Medicine