Helmholtz Gemeinschaft


Gene structure conservation aids similarity based gene prediction

Item Type:Article
Title:Gene structure conservation aids similarity based gene prediction
Creators Name:Meyer, I.M. and Durbin, R.
Abstract:One of the primary tasks in deciphering the functional contents of a newly sequenced genome is the identification of its protein coding genes. Existing computational methods for gene prediction include ab initio methods which use the DNA sequence itself as the only source of information, comparative methods using multiple genomic sequences, and similarity based methods which employ the cDNA or protein sequences of related genes to aid the gene prediction. We present here an algorithm implemented in a computer program called Projector which combines comparative and similarity approaches. Projector employs similarity information at the genomic DNA level by directly using known genes annotated on one DNA sequence to predict the corresponding related genes on another DNA sequence. It therefore makes explicit use of the conservation of the exon-intron structure between two related genes in addition to the similarity of their encoded amino acid sequences. We evaluate the performance of Projector by comparing it with the program Genewise on a test set of 491 pairs of independently confirmed mouse and human genes. It is more accurate than Genewise for genes whose proteins are <80% identical, and is suitable for use in a combined gene prediction system where other methods identify well conserved and non-conserved genes, and pseudogenes.
Keywords:Algorithms, Computational Biology, Conserved Sequence, Exons, Genes, Genomics, Initiator Codon, Introns, Nucleic Acid Sequence Homology, Open Reading Frames, Probability, Pseudogenes, Sensitivity and Specificity, Software, Terminator Codon, Animals, Mice
Source:Nucleic Acids Research
Publisher:Oxford University Press
Page Range:776-783
Date:4 February 2004
Official Publication:https://doi.org/10.1093/nar/gkh211
PubMed:View item in PubMed

Repository Staff Only: item control page

Open Access
MDC Library