Nucleic Acids Research

Multiple sequence alignment with hierarchical clustering

Journal article · 1988 · Cited by 5,419

✓ Free legal copy found

Preprint, hosted by National Institutes of Health (ncbi.nlm.nih.gov)

This is the authors’ own version from before peer review, so it may differ from the published paper.

Read it free at ncbi.nlm.nih.gov →

Abstract

An algorithm is presented for the multiple alignment of sequences, either proteins or nucleic acids, that is both accurate and easy to use on microcomputers. The approach is based on the conventional dynamic-programming method of pairwise alignment. Initially, a hierarchical clustering of the sequences is performed using the matrix of the pairwise alignment scores. The closest sequences are aligned creating groups of aligned sequences. Then close groups are aligned until all sequences are aligned in one group. The pairwise alignments included in the multiple alignment form a new matrix that is used to produce a hierarchical clustering. If it is different from the first one, iteration of the process can be performed. The method is illustrated by an example: a global alignment of 39 sequences of cytochrome c.

DOI: 10.1093/nar/16.22.10881 · Publisher: Oxford University Press (OUP)

Guides

Find another paper