An application of machine learning with feature selection to improve diagnosis and classification of neurodegenerative disorders

Josefa Díaz Álvarez,José L Risco-Martín,Jordi A Matias-Guiu,María Nieves Cabrera-Martín,José L Ayala

doi:10.1186/s12859-019-3027-7

Abstract

BackgroundThe analysis of health and medical data is crucial for improving the diagnosis precision, treatments and prevention. In this field, machine learning techniques play a key role. However, the amount of health data acquired from digital machines has high dimensionality and not all data acquired from digital machines are relevant for a particular disease. Primary Progressive Aphasia (PPA) is a neurodegenerative syndrome including several specific diseases, and it is a good model to implement machine learning analyses. In this work, we applied five feature selection algorithms to identify the set of relevant features from 18F-fluorodeoxyglucose positron emission tomography images of the main areas affected by PPA from patient records. On the other hand, we carried out classification and clustering algorithms before and after the feature selection process to contrast both results with those obtained in a previous work. We aimed to find the best classifier and the more relevant features from the WEKA tool to propose further a framework for automatic help on diagnosis. Dataset contains data from 150 FDG-PET imaging studies of 91 patients with a clinic prognosis of PPA, which were examined twice, and 28 controls. Our method comprises six different stages: (i) feature extraction, (ii) expertise knowledge supervision (iii) classification process, (iv) comparing classification results for feature selection, (v) clustering process after feature selection, and (vi) comparing clustering results with those obtained in a previous work.ResultsExperimental tests confirmed clustering results from a previous work. Although classification results for some algorithms are not decisive for reducing features precisely, Principal Components Analisys (PCA) results exhibited similar or even better performances when compared to those obtained with all features.ConclusionsAlthough reducing the dimensionality does not means a general improvement, the set of features is almost halved and results are better or quite similar. Finally, it is interesting how these results expose a finer grain classification of patients according to the neuroanatomy of their disease.

Highlights

The analysis of health and medical data is crucial for improving the diagnosis precision, treatments and prevention
Our study addresses the improvement of the Progressive Aphasia (PPA) diagnosis from Fluorodeoxyglucose positron emission tomography (FDG-Positron Emission Tomography (PET)) images applying machine learning techniques
This study confirms and reinforces the results obtained in a previous clinical work, where we explored the automatic classification of PPA patients and found out new subtypes of this disease that correlate with the clinical findings and better predict the clinical course

Summary

Introduction

The analysis of health and medical data is crucial for improving the diagnosis precision, treatments and prevention. Learning from data is one of the most successful fields applicable to many heterogeneous areas and disciplines like statistics, artificial intelligence, engineering, health, etc Digital machines such as Magnetic Resonance Imaging (MRI), Mass Spectrometry (MS) or Positron Emission Tomography (PET), among others, and new generation sensors, have found their way in biomedical systems. The amount of data has exponential growth, and the need to extract new knowledge for tackling diseases places bioinformatics as a priority research area In this regard, machine learning and big data methods have been applied to better understand and fight many diseases [3,4,5,6]

Objectives

Methods

Results

Discussion

Conclusion

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: BMC bioinformatics	Publication Date: Oct 11, 2019
Citations: 42	License type: open-access

R Discovery Prime

R Discovery Prime

An application of machine learning with feature selection to improve diagnosis and classification of neurodegenerative disorders

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: BMC bioinformatics

Lead the way for us

Similar Papers

Machine learning in pain research.
Jörn Lötsch ... Alfred Ultsch
PAIN | VOL. 159
Jörn Lötsch, et. al.Jörn Lötsch ... Alfred Ultsch
24 Nov 2017
PAIN | VOL. 159

Improved Prediction Accuracy with Reduced Feature Set Using Novel Binary Gravitational Search Optimization
Sankhadip Saha ... Dwaipayan Chakraborty
-
Sankhadip Saha, et. al.Sankhadip Saha ... Dwaipayan Chakraborty
01 Jan 2015
01 Jan 2015

Diagnosis of Alzheimer’s disease and frontotemporal dementia using FDG‐PET: Application of genetic algorithms
Vanesa Pytel ... Laura Hernández‐Lorenzo
Alzheimer's & Dementia: The Journal of the Alzheimer's Association | VOL. 17
Vanesa Pytel, et. al.Vanesa Pytel ... Laura Hernández‐Lorenzo
01 Dec 2021
Alzheimer's & Dementia: The Journal of the Alzheimer's Association | VOL. 17

Some issues on scalable feature selection
Huan Liu ... Rudy Setiono
Expert systems with applications | VOL. 15
Huan Liu, et. al.Huan Liu ... Rudy Setiono
01 Oct 1998
Expert systems with applications | VOL. 15

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

An application of machine learning with feature selection to improve diagnosis and classification of neurodegenerative disorders

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: BMC bioinformatics