SnowyOwl: accurate prediction of fungal genes by using RNA-Seq and homology information to select ab initio models.

Ian Reid*, Nick O'Toole, Omar Zabaneh, Reza Nourzadeh, Mike Dahdouli, Mostafa Abdellateef, Paul M. Gordon, Jung Soh, Greg Butler, Christoph Wilhelm Sensen, Adrian Tsang

*Korrespondierende/r Autor/-in für diese Arbeit

Publikation: Beitrag in einer FachzeitschriftArtikelBegutachtung

Abstract

Background
Locating the protein-coding genes in novel genomes is essential to understanding and exploiting the genomic information but it is still difficult to accurately predict all the genes. The recent availability of detailed information about transcript structure from high-throughput sequencing of messenger RNA (RNA-Seq) delineates many expressed genes and promises increased accuracy in gene prediction. Computational gene predictors have been intensively developed for and tested in well-studied animal genomes. Hundreds of fungal genomes are now or will soon be sequenced. The differences of fungal genomes from animal genomes and the phylogenetic sparsity of well-studied fungi call for gene-prediction tools tailored to them.
Results
SnowyOwl is a new gene prediction pipeline that uses RNA-Seq data to train and provide hints for the generation of Hidden Markov Model (HMM)-based gene predictions and to evaluate the resulting models. The pipeline has been developed and streamlined by comparing its predictions to manually curated gene models in three fungal genomes and validated against the high-quality gene annotation of Neurospora crassa; SnowyOwl predicted N. crassa genes with 83% sensitivity and 65% specificity. SnowyOwl gains sensitivity by repeatedly running the HMM gene predictor Augustus with varied input parameters and selectivity by choosing the models with best homology to known proteins and best agreement with the RNA-Seq data.
Conclusions
SnowyOwl efficiently uses RNA-Seq data to produce accurate gene models in both well-studied and novel fungal genomes. The source code for the SnowyOwl pipeline (in Python) and a web interface (in PHP) is freely available from http://sourceforge.net/projects/snowyowl/
Originalspracheenglisch
Seiten (von - bis)229-229
FachzeitschriftBMC Bioinformatics
Jahrgang15
Ausgabenummer1
DOIs
PublikationsstatusVeröffentlicht - 2014
Extern publiziertJa

Fields of Expertise

  • Human- & Biotechnology

Treatment code (Nähere Zuordnung)

  • Basic - Fundamental (Grundlagenforschung)

Fingerprint

Untersuchen Sie die Forschungsthemen von „SnowyOwl: accurate prediction of fungal genes by using RNA-Seq and homology information to select ab initio models.“. Zusammen bilden sie einen einzigartigen Fingerprint.

Dieses zitieren