首页 | 本学科首页   官方微博 | 高级检索  
     


5.8S-28S rRNA interaction and HMM-based ITS2 annotation
Authors:Alexander Keller, Tina Schleicher, J  rg Schultz, Tobias Mü  ller, Thomas Dandekar,Matthias Wolf
Affiliation:aDepartment of Bioinformatics, University of Würzburg, Biocenter, Am Hubland, 97074 Würzburg, Germany
Abstract:The internal transcribed spacer 2 (ITS2) of the nuclear ribosomal repeat unit is one of the most commonly applied phylogenetic markers. It is a fast evolving locus, which makes it appropriate for studies at low taxonomic levels, whereas its secondary structure is well conserved, and tree reconstructions are possible at higher taxonomic levels. However, annotation of start and end positions of the ITS2 differs markedly between studies. This is a severe shortcoming, as prediction of a correct secondary structure by standard ab initio folding programs requires accurate identification of the marker in question. Furthermore, the correct structure is essential for multiple sequence alignments based on individual structural features. The present study describes a new tool for the delimitation and identification of the ITS2. It is based on hidden Markov models (HMMs) and verifies annotations by comparison to a conserved structural motif in the 5.8S/28S rRNA regions. Our method was able to identify and delimit the ITS2 in more than 30 000 entries lacking start and end annotations in GenBank. Furthermore, 45 000 ITS2 sequences with a questionable annotation were re-annotated. Approximately 30 000 entries from the ITS2-DB, that uses a homology-based method for structure prediction, were re-annotated. We show that the method is able to correctly annotate an ITS2 as small as 58 nt from Giardia lamblia and an ITS2 as large as 1160 nt from humans. Thus, our method should be a valuable guide during the first and crucial step in any ITS2-based phylogenetic analysis: the delineation of the correct sequence. Sequences can be submitted to the following website for HMM-based ITS2 delineation: http://its2.bioapps.biozentrum.uni-wuerzburg.de.
Keywords:Hidden Markov models   Internal transcribed spacer 2   Non-coding RNA   Phylogenetics   Ribosomal RNA   Secondary structure
本文献已被 ScienceDirect 等数据库收录!
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号