Searching for convergence in phylogenetic Markov chain Monte Carlo
- PMID: 16857650
- DOI: 10.1080/10635150600812544
Searching for convergence in phylogenetic Markov chain Monte Carlo
Abstract
Markov chain Monte Carlo (MCMC) is a methodology that is gaining widespread use in the phylogenetics community and is central to phylogenetic software packages such as MrBayes. An important issue for users of MCMC methods is how to select appropriate values for adjustable parameters such as the length of the Markov chain or chains, the sampling density, the proposal mechanism, and, if Metropolis-coupled MCMC is being used, the number of heated chains and their temperatures. Although some parameter settings have been examined in detail in the literature, others are frequently chosen with more regard to computational time or personal experience with other data sets. Such choices may lead to inadequate sampling of tree space or an inefficient use of computational resources. We performed a detailed study of convergence and mixing for 70 randomly selected, putatively orthologous protein sets with different sizes and taxonomic compositions. Replicated runs from multiple random starting points permit a more rigorous assessment of convergence, and we developed two novel statistics, delta and epsilon, for this purpose. Although likelihood values invariably stabilized quickly, adequate sampling of the posterior distribution of tree topologies took considerably longer. Our results suggest that multimodality is common for data sets with 30 or more taxa and that this results in slow convergence and mixing. However, we also found that the pragmatic approach of combining data from several short, replicated runs into a "metachain" to estimate bipartition posterior probabilities provided good approximations, and that such estimates were no worse in approximating a reference posterior distribution than those obtained using a single long run of the same length as the metachain. Precision appears to be best when heated Markov chains have low temperatures, whereas chains with high temperatures appear to sample trees with high posterior probabilities only rarely.
Similar articles
-
Phylogenetic MCMC algorithms are misleading on mixtures of trees.Science. 2005 Sep 30;309(5744):2207-9. doi: 10.1126/science.1115493. Science. 2005. PMID: 16195459
-
Guided tree topology proposals for Bayesian phylogenetic inference.Syst Biol. 2012 Jan;61(1):1-11. doi: 10.1093/sysbio/syr074. Epub 2011 Aug 9. Syst Biol. 2012. PMID: 21828081
-
An examination of the monophyly of morning glory taxa using Bayesian phylogenetic inference.Syst Biol. 2002 Oct;51(5):740-53. doi: 10.1080/10635150290102401. Syst Biol. 2002. PMID: 12396588
-
Finding starting points for Markov chain Monte Carlo analysis of genetic data from large and complex pedigrees.Genet Epidemiol. 2003 Jul;25(1):14-24. doi: 10.1002/gepi.10243. Genet Epidemiol. 2003. PMID: 12813723 Review.
-
Potential applications and pitfalls of Bayesian inference of phylogeny.Syst Biol. 2002 Oct;51(5):673-88. doi: 10.1080/10635150290102366. Syst Biol. 2002. PMID: 12396583 Review.
Cited by
-
Quantifying MCMC exploration of phylogenetic tree space.Syst Biol. 2015 May;64(3):472-91. doi: 10.1093/sysbio/syv006. Epub 2015 Jan 27. Syst Biol. 2015. PMID: 25631175 Free PMC article.
-
Protein evolution by molecular tinkering: diversification of the nuclear receptor superfamily from a ligand-dependent ancestor.PLoS Biol. 2010 Oct 5;8(10):e1000497. doi: 10.1371/journal.pbio.1000497. PLoS Biol. 2010. PMID: 20957188 Free PMC article.
-
Evolution of the parasitic wasp subfamily Rogadinae (Braconidae): phylogeny and evolution of lepidopteran host ranges and mummy characteristics.BMC Evol Biol. 2008 Dec 4;8:329. doi: 10.1186/1471-2148-8-329. BMC Evol Biol. 2008. PMID: 19055825 Free PMC article.
-
Evolution of general transcription factors.J Mol Evol. 2013 Feb;76(1-2):28-47. doi: 10.1007/s00239-012-9535-y. Epub 2012 Dec 11. J Mol Evol. 2013. PMID: 23229069
-
Bipartite Network Analysis of Gene Sharings in the Microbial World.Mol Biol Evol. 2018 Apr 1;35(4):899-913. doi: 10.1093/molbev/msy001. Mol Biol Evol. 2018. PMID: 29346651 Free PMC article.
Publication types
MeSH terms
LinkOut - more resources
Full Text Sources