Scalable and highly parallel implementation of Smith-Waterman on graphics processing unit using CUDA

Ali Akoglu,Gregory M Striemer

doi:10.1007/s10586-009-0089-8

Abstract

Program development environments have enabled graphics processing units (GPUs) to become an attractive high performance computing platform for the scientific community. A commonly posed problem in computational biology is protein database searching for functional similarities. The most accurate algorithm for sequence alignments is Smith-Waterman (SW). However, due to its computational complexity and rapidly increasing database sizes, the process becomes more and more time consuming making cluster based systems more desirable. Therefore, scalable and highly parallel methods are necessary to make SW a viable solution for life science researchers. In this paper we evaluate how SW fits onto the target GPU architecture by exploring ways to map the program architecture on the processor architecture. We develop new techniques to reduce the memory footprint of the application while exploiting the memory hierarchy of the GPU. With this implementation, GSW, we overcome the on chip memory size constraint, achieving 23× speedup compared to a serial implementation. Results show that as the query length increases our speedup almost stays stable indicating the solid scalability of our approach. Additionally this is a first of a kind implementation which purely runs on the GPU instead of a CPU-GPU integrated environment, making our design suitable for porting onto a cluster of GPUs.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Scalable and highly parallel implementation of Smith-Waterman on graphics processing unit using CUDA

Abstract

Talk to us

Similar Papers

More From: Cluster Computing

Lead the way for us

Journal: Cluster Computing	Publication Date: Jun 11, 2009
Citations: 23

Similar Papers

GEYSER: 3D thermo-hydrodynamic reactive transport numerical simulator including porosity and permeability evolution using GPU clusters
Reza Sohrabi ... Samuel Omlin
Computational Geosciences | VOL. 23
Reza Sohrabi, et. al.Reza Sohrabi ... Samuel Omlin
13 Sep 2019
Computational Geosciences | VOL. 23

Fast distributed large-pixel-count hologram computation using a GPU cluster
Yuechao Pan ... Xuewu Xu
Applied Optics | VOL. 52
Yuechao Pan, et. al.Yuechao Pan ... Xuewu Xu
09 Sep 2013
Applied Optics | VOL. 52

Overlapping computation and communication of three-dimensional FDTD on a GPU cluster
Ki-Hwan Kim ... Q-Han Park
Computer Physics Communications | VOL. 183
Ki-Hwan Kim, et. al.Ki-Hwan Kim ... Q-Han Park
15 Jun 2012
Computer Physics Communications | VOL. 183

HASEonGPU—An adaptive, load-balanced MPI/GPU-code for calculating the amplified spontaneous emission in high power laser media
C.H.J Eckert ... D Albach
Computer Physics Communications | VOL. 207
C.H.J Eckert, et. al.C.H.J Eckert ... D Albach
30 May 2016
Computer Physics Communications | VOL. 207

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Scalable and highly parallel implementation of Smith-Waterman on graphics processing unit using CUDA

Abstract

Talk to us

Similar Papers

More From: Cluster Computing