Back to Search
Start Over
PASTA: Ultra-Large Multiple Sequence Alignment for Nucleotide and Amino-Acid Sequences.
- Source :
-
Journal of computational biology : a journal of computational molecular cell biology [J Comput Biol] 2015 May; Vol. 22 (5), pp. 377-86. Date of Electronic Publication: 2014 Dec 30. - Publication Year :
- 2015
-
Abstract
- We introduce PASTA, a new multiple sequence alignment algorithm. PASTA uses a new technique to produce an alignment given a guide tree that enables it to be both highly scalable and very accurate. We present a study on biological and simulated data with up to 200,000 sequences, showing that PASTA produces highly accurate alignments, improving on the accuracy and scalability of the leading alignment methods (including SATé). We also show that trees estimated on PASTA alignments are highly accurate--slightly better than SATé trees, but with substantial improvements relative to other methods. Finally, PASTA is faster than SATé, highly parallelizable, and requires relatively little memory.
- Subjects :
- Amino Acid Sequence
Bacteria classification
Base Sequence
Computer Simulation
Datasets as Topic
Evolution, Molecular
Metagenomics methods
Molecular Sequence Data
Phylogeny
Plants classification
Sequence Alignment methods
Algorithms
Bacteria genetics
Metagenomics statistics & numerical data
Plants genetics
Sequence Alignment statistics & numerical data
Software
Subjects
Details
- Language :
- English
- ISSN :
- 1557-8666
- Volume :
- 22
- Issue :
- 5
- Database :
- MEDLINE
- Journal :
- Journal of computational biology : a journal of computational molecular cell biology
- Publication Type :
- Academic Journal
- Accession number :
- 25549288
- Full Text :
- https://doi.org/10.1089/cmb.2014.0156