Home LiteratureArticle Details
PMID: 8108430 Published · ppublish English Journal Article

Reconstructing evolutionary trees from DNA and protein sequences: paralinear distances.

Lake JA

Abstract

The reconstruction of phylogenetic trees from DNA and protein sequences is confounded by unequal rate effects. These effects can group rapidly evolving taxa with other rapidly evolving taxa, whether or not they are genealogically related. All algorithms are sensitive to these effects whenever the assumptions on which they are based are not met. The algorithm presented here, called paralinear distances, is valid for a much broader class of substitution processes than previous algorithms and is accordingly less affected by unequal rate effects. It may be used with all nucleic acid, protein, or other sequences, provided that their evolution may be modeled as a succession of Markov processes. The properties of the method have been proven both analytically and by computer simulations. Like all other methods, paralinear distances can fail when sequences are misaligned or when site-to-site sequence variation of rates is extensive. To examine the usefulness of paralinear distances, the "origin of the eukaryotes" has been investigated by the analysis of elongation factor Tu sequences with a variety of sequence alignments. It has been found that the order in which sequences are pairwise aligned strongly determines the topology which is reconstructed by paralinear distances (as it does for all other reconstruction methods tested). When the parts of the alignment that are unaffected by alignment order are analyzed, paralinear distances strongly select the eocyte topology. This provides evidence that the eocyte prokaryotes are the closest prokaryotic relatives of the eukaryotes.

MeSH Terms
Algorithms Archaea/genetics Base Sequence Biological Evolution Computer Simulation DNA/genetics Eukaryotic Cells Halobacteriales/genetics Markov Chains Models, Genetic Molecular Sequence Data Proteins/genetics Sequence Alignment/methods
Chemicals
Proteins DNA
Authors & Affiliations
1 authors, click to expand affiliations / ORCID
Lake J A
Molecular Biology Institute, University of California, Los Angeles 90024.
References (16)
16 references, click to expand
  1. The sequence of the gene encoding elongation factor Tu from Chlamydia trachomatis compared with those of other organisms.
    Gene. 1992 Oct 12;120(1):33-41 PMID: 1398121
  2. Evidence that eukaryotes and eocyte prokaryotes are immediate relatives.
    Science. 1992 Jul 3;257(5066):74-6 PMID: 1621096
  3. Nucleotide sequence of the gene for elongation factor EF-1 alpha from the extreme thermophilic archaebacterium Thermococcus celer.
    Nucleic Acids Res. 1990 Jul 11;18(13):3989 PMID: 2115672
  4. The order of sequence alignment can bias the selection of tree topology.
    Mol Biol Evol. 1991 May;8(3):378-85 PMID: 2072863
  5. A rate-independent technique for analysis of nucleic acid sequences: evolutionary parsimony.
    Mol Biol Evol. 1987 Mar;4(2):167-91 PMID: 3447007
  6. Origin of the eukaryotic nucleus determined by rate-invariant analysis of rRNA sequences.
    Nature. 1988 Jan 14;331(6152):184-6 PMID: 3340165
  7. Confidence in evolutionary trees from biological sequence data.
    Nature. 1993 Jul 29;364(6436):440-2 PMID: 8332213
  8. Nucleotide sequence of a DNA region comprising the gene for elongation factor 1 alpha (EF-1 alpha) from the ultrathermophilic archaeote Pyrococcus woesei: phylogenetic implications.
    J Mol Evol. 1991 Oct;33(4):332-42 PMID: 1723106
  9. Optimal sequence alignments.
    Proc Natl Acad Sci U S A. 1983 Mar;80(5):1382-6 PMID: 16593289
  10. Functional implications related to the gene structure of the elongation factor EF-Tu from Halobacterium marismortui.
    Nucleic Acids Res. 1990 Feb 11;18(3):507-11 PMID: 2155402
  11. The nucleotide sequence of the cloned tufA gene of Escherichia coli.
    Gene. 1980 Dec;12(1-2):25-31 PMID: 7011903
  12. Asynchronous distance between homologous DNA sequences.
    Biometrics. 1987 Jun;43(2):261-76 PMID: 3607200
  13. Polypeptide chain elongation factor 1 alpha (EF-1 alpha) from yeast: nucleotide sequence of one of the two genes for EF-1 alpha from Saccharomyces cerevisiae.
    EMBO J. 1984 Aug;3(8):1825-30 PMID: 6383821
  14. Establishing homologies in protein sequences.
    Methods Enzymol. 1983;91:524-45 PMID: 6855599
  15. Dynamic programming algorithms for biological sequence comparison.
    Methods Enzymol. 1992;210:575-601 PMID: 1584052
  16. The evolution of the globin family genes: concordance of stochastic and augmented maximum parsimony genetic distances for alpha hemoglobin, beta hemoglobin and myoglobin phylogenies.
    J Mol Biol. 1976 Jul 25;105(1):39-74 PMID: 994186
Article Info
Journal
Proceedings of the National Academy of Sciences of the United States of America
Abbr.
Proc Natl Acad Sci U S A
ISSN
0027-8424
Published
1994-02-15
Pages
1455-9
Language
English
Region
United States
NLM ID
7505876
PMCID
PMC43178
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: [email protected]