Sparse Neighbor Joining: rapid phylogenetic inference using a sparse distance matrix.

Semih Kurt,Alexandre Bouchard-Côté,Jens Lagergren

doi:10.1093/bioinformatics/btae701

Abstract

Phylogenetic reconstruction is a fundamental problem in computational biology. The Neighbor Joining (NJ) algorithm offers an efficient distance-based solution to this problem, which often serves as the foundation for more advanced statistical methods. Despite prior efforts to enhance the speed of NJ, the computation of the n 2 entries of the distance matrix, where n is the number of phylogenetic tree leaves, continues to pose a limitation in scaling NJ to larger datasets. In this work, we propose a new algorithm which does not require computing a dense distance matrix. Instead, it dynamically determines a sparse set of at most O(n log n) distance matrix entries to be computed in its basic version, and up to O(n log 2n) entries in an enhanced version. We show by experiments that this approach reduces the execution time of NJ for large datasets, with a trade-off in accuracy. Sparse Neighbor Joining is implemented in Python and freely available at https://github.com/kurtsemih/SNJ. Supplementary data are available at Bioinformatics online.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Sparse Neighbor Joining: rapid phylogenetic inference using a sparse distance matrix.

Abstract

Talk to us

Similar Papers

More From: Bioinformatics (Oxford, England)

Lead the way for us

Similar Papers

Improved prediction of post-translational modification crosstalk within proteins using DeepPCT.
Yu-Xiang Huang ... Rong Liu
Bioinformatics (Oxford, England) | VOL. -
Yu-Xiang Huang, et. al.Yu-Xiang Huang ... Rong Liu
21 Nov 2024
Bioinformatics (Oxford, England) | VOL. -

Sparse Neighbor Joining: rapid phylogenetic inference using a sparse distance matrix.
Semih Kurt ... Jens Lagergren
Bioinformatics (Oxford, England) | VOL. -
Semih Kurt, et. al.Semih Kurt ... Jens Lagergren
21 Nov 2024
Bioinformatics (Oxford, England) | VOL. -

OneSC: A computational platform for recapitulating cell state transitions.
Da Peng ... Patrick Cahan
Bioinformatics (Oxford, England) | VOL. -
Da Peng, et. al.Da Peng ... Patrick Cahan
21 Nov 2024
Bioinformatics (Oxford, England) | VOL. -

Accurate and Transferable Drug-Target Interaction Prediction with DrugLAMP.
Zhengchao Luo ... Jinzhuo Wang
Bioinformatics (Oxford, England) | VOL. -
Zhengchao Luo, et. al.Zhengchao Luo ... Jinzhuo Wang
21 Nov 2024
Bioinformatics (Oxford, England) | VOL. -

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Sparse Neighbor Joining: rapid phylogenetic inference using a sparse distance matrix.

Abstract

Talk to us

Similar Papers

More From: Bioinformatics (Oxford, England)