Publications

Detailed Information

Block Aligner: an adaptive SIMD-accelerated aligner for sequences and position-specific scoring matrices

DC Field Value Language
dc.contributor.authorLiu, Daniel-
dc.contributor.authorSteinegger, Martin-
dc.date.accessioned2024-05-16T01:26:00Z-
dc.date.available2024-05-16T01:26:00Z-
dc.date.created2023-09-13-
dc.date.created2023-09-13-
dc.date.issued2023-08-
dc.identifier.citationBioinformatics, Vol.39 No.8, p. 487-
dc.identifier.issn1367-4803-
dc.identifier.urihttps://hdl.handle.net/10371/202499-
dc.description.abstractMOTIVATION: Efficiently aligning sequences is a fundamental problem in bioinformatics. Many recent algorithms for computing alignments through Smith-Waterman-Gotoh dynamic programming (DP) exploit Single Instruction Multiple Data (SIMD) operations on modern CPUs for speed. However, these advances have largely ignored difficulties associated with efficiently handling complex scoring matrices or large gaps (insertions or deletions). RESULTS: We propose a new SIMD-accelerated algorithm called Block Aligner for aligning nucleotide and protein sequences against other sequences or position-specific scoring matrices. We introduce a new paradigm that uses blocks in the DP matrix that greedily shift, grow, and shrink. This approach allows regions of the DP matrix to be adaptively computed. Our algorithm reaches over 5-10 times faster than some previous methods while incurring an error rate of less than 3% on protein and long read datasets, despite large gaps and low sequence identities. AVAILABILITY AND IMPLEMENTATION: Our algorithm is implemented for global, local, and X-drop alignments. It is available as a Rust library (with C bindings) at https://github.com/Daniel-Liu-c0deb0t/block-aligner.-
dc.language영어-
dc.publisherOxford University Press-
dc.titleBlock Aligner: an adaptive SIMD-accelerated aligner for sequences and position-specific scoring matrices-
dc.typeArticle-
dc.identifier.doi10.1093/bioinformatics/btad487-
dc.citation.journaltitleBioinformatics-
dc.identifier.wosid001187250700001-
dc.identifier.scopusid2-s2.0-85168810132-
dc.citation.number8-
dc.citation.startpage487-
dc.citation.volume39-
dc.description.isOpenAccessY-
dc.contributor.affiliatedAuthorSteinegger, Martin-
dc.type.docTypeArticle-
dc.description.journalClass1-
dc.subject.keywordPlusALGORITHM-
dc.subject.keywordPlusALIGNMENT-
dc.subject.keywordPlusPARALLEL-
dc.subject.keywordPlusSEARCH-
dc.subject.keywordPlusSPEED-
Appears in Collections:
Files in This Item:
There are no files associated with this item.

Related Researcher

  • College of Natural Sciences
  • School of Biological Sciences
Research Area Development of algorithms to search, cluster and assemble sequence data, Metagenomic analysis, Pathogen detection in sequencing data

Altmetrics

Item View & Download Count

  • mendeley

Items in S-Space are protected by copyright, with all rights reserved, unless otherwise indicated.

Share