Abstract
Improvements in technology have triggered the production of big data. Within this scope, enormous amounts of biological data have been generated. A number of analysis methods have been developed to access the information contained in biological data. DNA sequence analysis has drawn particular attention in recent years. As an alternative to alignment-based sequence comparison methods that have high computational costs, alignment-free comparison methods have emerged. These m ethods c an c alculate sequence s imilarity by applying different dimensions of numerical characterizations. In this paper, we propose a novel alignment-free DNA sequence analysis method based on a feature extraction strategy. The method utilizes numerical characterization and is implemented by calculating mean distance of the transitions, mean distance of the nucleotide duplications, and the base frequencies. The method then measures the similarity between 7-dimensional vectors that are obtained through feature extraction. Using this approach, we conducted a sequence similarity analysis of two different DNA sequence datasets of different lengths to demonstrate the effectiveness of the method. The proposed method shows that a simple and successful feature vector can be obtained when DNA sequences having many properties are used in combination with appropriate and effective descriptors. With this strategy, reasonable results were obtained with a low computational cost.
Talk to us
Join us for a 30 min session where you can share your feedback and ask us any queries you have
More From: Sigma Journal of Engineering and Natural Sciences – Sigma Mühendislik ve Fen Bilimleri Dergisi
Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.