Progressive multiple sequence alignments from triplets

被引:17
|
作者
Kruspe, Matthias
Stadler, Peter F.
机构
[1] Univ Leipzig, Bioinformat Grp, Dept Comp Sci, D-04107 Leipzig, Germany
[2] Univ Leipzig, Bioinformat Grp, Interdisciplinary Ctr Bioinformat, D-04107 Leipzig, Germany
[3] Fraunhofer Inst Zelltherapie & Immunol IZI, D-04103 Leipzig, Germany
[4] Univ Vienna, Inst Theoret Chem, A-1090 Vienna, Austria
[5] Santa Fe Inst, Santa Fe, NM 87501 USA
关键词
D O I
10.1186/1471-2105-8-254
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Background: The quality of progressive sequence alignments strongly depends on the accuracy of the individual pairwise alignment steps since gaps that are introduced at one step cannot be removed at later aggregation steps. Adjacent insertions and deletions necessarily appear in arbitrary order in pairwise alignments and hence form an unavoidable source of errors. Research: Here we present a modified variant of progressive sequence alignments that addresses both issues. Instead of pairwise alignments we use exact dynamic programming to align sequence or profile triples. This avoids a large fractions of the ambiguities arising in pairwise alignments. In the subsequent aggregation steps we follow the logic of the Neighbor- Net algorithm, which constructs a phylogenetic network by step- wisely replacing triples by pairs instead of combining pairs to singletons. To this end the three- way alignments are subdivided into two partial alignments, at which stage all- gap columns are naturally removed. This alleviates the '' once a gap, always a gap '' problem of progressive alignment procedures. Conclusion: The three- way Neighbor- Net based alignment program aln3nn is shown to compare favorably on both protein sequences and nucleic acids sequences to other progressive alignment tools. In the latter case one easily can include scoring terms that consider secondary structure features. Overall, the quality of resulting alignments in general exceeds that of clustalw or other multiple alignments tools even though our software does not included heuristics for context dependent ( mis) match scores.
引用
收藏
页数:12
相关论文
共 50 条
  • [41] BUILDING MULTIPLE ALIGNMENTS FROM PAIRWISE ALIGNMENTS
    MILLER, W
    COMPUTER APPLICATIONS IN THE BIOSCIENCES, 1993, 9 (02): : 169 - 176
  • [42] Gleaning structural and functional information from correlations in protein multiple sequence alignments
    Neuwald, Andrew F.
    CURRENT OPINION IN STRUCTURAL BIOLOGY, 2016, 38 : 1 - 8
  • [43] ESTIMATING POLYPEPTIDE ALPHA-CARBON DISTANCES FROM MULTIPLE SEQUENCE ALIGNMENTS
    ASZODI, A
    TAYLOR, WR
    JOURNAL OF MATHEMATICAL CHEMISTRY, 1995, 17 (2-3) : 167 - 184
  • [44] Easy method to predict solvent accessibility from multiple protein sequence alignments
    Pascarella, S
    De Persio, R
    Bossa, F
    Argos, P
    PROTEINS-STRUCTURE FUNCTION AND GENETICS, 1998, 32 (02): : 190 - 199
  • [45] Natural amino acid classification scheme derived from multiple sequence alignments
    Fauman, EB
    ABSTRACTS OF PAPERS OF THE AMERICAN CHEMICAL SOCIETY, 2005, 229 : U796 - U796
  • [46] Erratum to: Sequence Diversity Diagram for comparative analysis of multiple sequence alignments
    Ryo Sakai
    Jan Aerts
    BMC Proceedings, 8 (Suppl 2)
  • [47] Noisy:: Identification of problematic columns in multiple sequence alignments
    Dress, Andreas W. M.
    Flamm, Christoph
    Fritzsch, Guido
    Gruenewald, Stefan
    Kruspe, Matthias
    Prohaska, Sonja J.
    Stadler, Peter F.
    ALGORITHMS FOR MOLECULAR BIOLOGY, 2008, 3 (1)
  • [48] Refining multiple sequence alignments with conserved core regions
    Chakrabarti, Saikat
    Lanczycki, Christopher J.
    Panchenko, Anna R.
    Przytycka, Teresa M.
    Thiessen, Paul A.
    Bryant, Stephen H.
    NUCLEIC ACIDS RESEARCH, 2006, 34 (09) : 2598 - 2606
  • [49] Erratum to: State of the art: refinement of multiple sequence alignments
    Saikat Chakrabarti
    Christopher J Lanczycki
    Anna R Panchenko
    Teresa M Przytycka
    Paul A Thiessen
    Stephen H Bryant
    BMC Bioinformatics, 11
  • [50] Optimization with Genetic Algorithm for Outcome of Multiple Sequence Alignments
    Li, Hongbin
    Zhang, Meile
    8TH INTERNATIONAL CONFERENCE ON BIOINFORMATICS AND BIOMEDICAL ENGINEERING (ICBBE 2014), 2014, : 25 - 30