Correcting for Sequencing Error in Maximum Likelihood Phylogeny Inference

被引:8
|
作者
Kuhner, Mary K. [1 ]
McGill, James [1 ]
机构
[1] Univ Washington, Dept Genome Sci, Seattle, WA 98195 USA
来源
G3-GENES GENOMES GENETICS | 2014年 / 4卷 / 12期
基金
美国国家科学基金会;
关键词
sequencing error; phylogeny inference; maximum likelihood; TREES;
D O I
10.1534/g3.114.014365
中图分类号
Q3 [遗传学];
学科分类号
071007 ; 090102 ;
摘要
Accurate phylogenies are critical to taxonomy as well as studies of speciation processes and other evolutionary patterns. Accurate branch lengths in phylogenies are critical for dating and rate measurements. Such accuracy may be jeopardized by unacknowledged sequencing error. We use simulated data to test a correction for DNA sequencing error in maximum likelihood phylogeny inference. Over a wide range of data polymorphism and true error rate, we found that correcting for sequencing error improves recovery of the branch lengths, even if the assumed error rate is up to twice the true error rate. Low error rates have little effect on recovery of the topology. When error is high, correction improves topological inference; however, when error is extremely high, using an assumed error rate greater than the true error rate leads to poor recovery of both topology and branch lengths. The error correction approach tested here was proposed in 2004 but has not been widely used, perhaps because researchers do not want to commit to an estimate of the error rate. This study shows that correction with an approximate error rate is generally preferable to ignoring the issue.
引用
收藏
页码:2544 / 2551
页数:8
相关论文
共 50 条
  • [1] Genetic programming for maximum-likelihood phylogeny inference
    Lv, H. Y.
    Zhou, C. G.
    Zhou, J. B.
    COMPUTATIONAL METHODS, PTS 1 AND 2, 2006, : 1255 - +
  • [2] Parallel inference of a 10.000-taxon phylogeny with maximum likelihood
    Stamatakis, A
    Ludwig, T
    Meier, H
    EURO-PAR 2004 PARALLEL PROCESSING, PROCEEDINGS, 2004, 3149 : 997 - 1004
  • [3] MAXIMUM-LIKELIHOOD INFERENCE OF PROTEIN PHYLOGENY AND THE ORIGIN OF CHLOROPLASTS
    KISHINO, H
    MIYATA, T
    HASEGAWA, M
    JOURNAL OF MOLECULAR EVOLUTION, 1990, 31 (02) : 151 - 160
  • [4] Genetic algorithms and parallel processing in maximum-likelihood phylogeny inference
    Brauer, MJ
    Holder, MT
    Dries, LA
    Zwickl, DJ
    Lewis, PO
    Hillis, DM
    MOLECULAR BIOLOGY AND EVOLUTION, 2002, 19 (10) : 1717 - 1726
  • [5] SUCCESS OF MAXIMUM-LIKELIHOOD PHYLOGENY INFERENCE IN THE 4-TAXON CASE
    GAUT, BS
    LEWIS, PO
    MOLECULAR BIOLOGY AND EVOLUTION, 1995, 12 (01) : 152 - 162
  • [6] Embedded computation of maximum-likelihood phylogeny inference using platform FPGA
    Mak, TST
    Lam, KP
    2004 IEEE COMPUTATIONAL SYSTEMS BIOINFORMATICS CONFERENCE, PROCEEDINGS, 2004, : 512 - 514
  • [7] Correcting energy balance error in heat exchanger data by maximum likelihood method
    Park, Young-Gil
    APPLIED THERMAL ENGINEERING, 2018, 131 : 311 - 319
  • [8] CNETML: maximum likelihood inference of phylogeny from copy number profiles of multiple samples
    Lu, Bingxin
    Curtius, Kit
    Graham, Trevor A.
    Yang, Ziheng
    Barnes, Chris P.
    GENOME BIOLOGY, 2023, 24 (01)
  • [9] CNETML: maximum likelihood inference of phylogeny from copy number profiles of multiple samples
    Bingxin Lu
    Kit Curtius
    Trevor A. Graham
    Ziheng Yang
    Chris P. Barnes
    Genome Biology, 24
  • [10] A genetic algorithm for maximum-likelihood phylogeny inference using nucleotide sequence data
    Lewis, PO
    MOLECULAR BIOLOGY AND EVOLUTION, 1998, 15 (03) : 277 - 283