Home LiteratureArticle Details
PMID: 17553837 Published · ppublish English Journal Article Research Support, Non-U.S. Gov't

MultiPhyl: a high-throughput phylogenomics webserver using distributed computing.

Nucleic acids research ·Vol. 35 ·No. Web Server issue ·2007-07-00 ·Pages W33-7

Keane TM, Naughton TJ, McInerney JO

Abstract

With the number of fully sequenced genomes increasing steadily, there is greater interest in performing large-scale phylogenomic analyses from large numbers of individual gene families. Maximum likelihood (ML) has been shown repeatedly to be one of the most accurate methods for phylogenetic construction. Recently, there have been a number of algorithmic improvements in maximum-likelihood-based tree search methods. However, it can still take a long time to analyse the evolutionary history of many gene families using a single computer. Distributed computing refers to a method of combining the computing power of multiple computers in order to perform some larger overall calculation. In this article, we present the first high-throughput implementation of a distributed phylogenetics platform, MultiPhyl, capable of using the idle computational resources of many heterogeneous non-dedicated machines to form a phylogenetics supercomputer. MultiPhyl allows a user to upload hundreds or thousands of amino acid or nucleotide alignments simultaneously and perform computationally intensive tasks such as model selection, tree searching and bootstrapping of each of the alignments using many desktop machines. The program implements a set of 88 amino acid models and 56 nucleotide maximum likelihood models and a variety of statistical methods for choosing between alternative models. A MultiPhyl webserver is available for public use at: http://www.cs.nuim.ie/distributed/multiphyl.php.

MeSH Terms
Algorithms Animals Computational Biology/methods Computer Simulation Computers Computing Methodologies Databases, Genetic Genomics/methods Humans Internet Likelihood Functions Phylogeny Sequence Alignment Software User-Computer Interface
Authors & Affiliations
3 authors, click to expand affiliations / ORCID
Keane Thomas M
Pathogen Sequencing Unit, Wellcome Trust Sanger Institute, Wellcome Trust Genome Campus, Hinxton, CB10 1SA Hinxton, UK. tkeane@cs.nuim.ie
Naughton Thomas J
McInerney James O
References (16)
16 references, click to expand
  1. Genetic algorithms and parallel processing in maximum-likelihood phylogeny inference.
    Mol Biol Evol. 2002 Oct;19(10):1717-26 PMID: 12270898
  2. RAxML-III: a fast program for maximum likelihood-based inference of large phylogenetic trees.
    Bioinformatics. 2005 Feb 15;21(4):456-63 PMID: 15608047
  3. PAL: an object-oriented programming library for molecular evolution and phylogenetics.
    Bioinformatics. 2001 Jul;17(7):662-3 PMID: 11448888
  4. TREE-PUZZLE: maximum likelihood phylogenetic analysis using quartets and parallel computing.
    Bioinformatics. 2002 Mar;18(3):502-4 PMID: 11934758
  5. A simple, fast, and accurate algorithm to estimate large phylogenies by maximum likelihood.
    Syst Biol. 2003 Oct;52(5):696-704 PMID: 14530136
  6. IQPNNI: moving fast through tree space and stopping in time.
    Mol Biol Evol. 2004 Aug;21(8):1565-71 PMID: 15163768
  7. The general stochastic model of nucleotide substitution.
    J Theor Biol. 1990 Feb 22;142(4):485-501 PMID: 2338834
  8. fastDNAmL: a tool for construction of phylogenetic trees of DNA sequences using maximum likelihood.
    Comput Appl Biosci. 1994 Feb;10(1):41-8 PMID: 8193955
  9. DPRml: distributed phylogeny reconstruction by maximum likelihood.
    Bioinformatics. 2005 Apr 1;21(7):969-74 PMID: 15513992
  10. DSEARCH: sensitive database searching using distributed computing.
    Bioinformatics. 2005 Apr 15;21(8):1705-6 PMID: 15564297
  11. The Opisthokonta and the Ecdysozoa may not be clades: stronger support for the grouping of plant and animal than for animal and fungi and stronger support for the Coelomata than Ecdysozoa.
    Mol Biol Evol. 2005 May;22(5):1175-84 PMID: 15703245
  12. Phylogenomics and the reconstruction of the tree of life.
    Nat Rev Genet. 2005 May;6(5):361-75 PMID: 15861208
  13. Distributed computing. Grassroots supercomputing.
    Science. 2005 May 6;308(5723):810 PMID: 15879205
  14. Assessment of methods for amino acid matrix selection and their use on empirical data shows that ad hoc assumptions for choice of matrix are not justified.
    BMC Evol Biol. 2006;6:29 PMID: 16563161
  15. A fungal phylogeny based on 42 complete genomes derived from supertree and combined gene analysis.
    BMC Evol Biol. 2006;6:99 PMID: 17121679
  16. MrBayes 3: Bayesian phylogenetic inference under mixed models.
    Bioinformatics. 2003 Aug 12;19(12):1572-4 PMID: 12912839
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
1362-4962
Published
2007-07-00
Epub
2007-00-06
Pages
W33-7
Language
English
Region
England
NLM ID
0411011
PMCID
PMC1933173
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com