A Flow Procedure for Linearization of Genome Sequence Graphs.

David Haussler,Maria Zueva,Maciej Smuga-Otto,Sergei Nikitin,Adam M Novak,Dmitrii Miagkov,Benedict Paten,Jordan M Eizenga

doi:10.1089/cmb.2017.0248

Abstract

Efforts to incorporate human genetic variation into the reference human genome have converged on the idea of a graph representation of genetic variation within a species, a genome sequence graph. A sequence graph represents a set of individual haploid reference genomes as paths in a single graph. When that set of reference genomes is sufficiently diverse, the sequence graph implicitly contains all frequent human genetic variations, including translocations, inversions, deletions, and insertions. In representing a set of genomes as a sequence graph, one encounters certain challenges. One of the most important is the problem of graph linearization, essential both for efficiency of storage and access, and for natural graph visualization and compatibility with other tools. The goal of graph linearization is to order nodes of the graph in such a way that operations such as access, traversal, and visualization are as efficient and effective as possible. A new algorithm for the linearization of sequence graphs, called the flow procedure (FP), is proposed in this article. Comparative experimental evaluation of the FP against other algorithms shows that it outperforms its rivals in the metrics most relevant to sequence graphs.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A Flow Procedure for Linearization of Genome Sequence Graphs.

Abstract

Talk to us

Similar Papers

More From: Journal of Computational Biology

Lead the way for us

Journal: Journal of Computational Biology	Publication Date: May 24, 2018
Citations: 3

Similar Papers

A Flow Procedure for the Linearization of Genome Sequence Graphs
David Haussler ... Maciej Smuga-Otto
-
David Haussler, et. al.David Haussler ... Maciej Smuga-Otto
01 Jan 2017
01 Jan 2017

AAPA Statement on Race and Racism.
Agustín Fuentes ... Tina Lasisi
American Journal of Physical Anthropology | VOL. 169
Agustín Fuentes, et. al.Agustín Fuentes ... Tina Lasisi
14 Jun 2019
American Journal of Physical Anthropology | VOL. 169

K2Mem: Discovering Discriminative K-mers From Sequencing Data for Metagenomic Reads Classification.
Davide Storato ... Matteo Comin
IEEE/ACM Transactions on Computational Biology and Bioinformatics | VOL. 19
Davide Storato, et. al.Davide Storato ... Matteo Comin
01 Jan 2021
IEEE/ACM Transactions on Computational Biology and Bioinformatics | VOL. 19

Improving Metagenomic Classification Using Discriminative k-mers from Sequencing Data
Davide Storato ... Matteo Comin
-
Davide Storato, et. al.Davide Storato ... Matteo Comin
01 Jan 2020
01 Jan 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Flow Procedure for Linearization of Genome Sequence Graphs.

Abstract

Talk to us

Similar Papers

More From: Journal of Computational Biology