Identifying Potential TRNA Genes in Genomic DNA Sequences
Overview
Molecular Biology
Authors
Affiliations
We have developed an algorithm that automatically and reproducibly identifies potential tRNA genes in genomic DNA sequences, and we present a general strategy for testing the sensitivity of such algorithms. This algorithm is useful for the flagging and characterization of long genomic sequences that have not been experimentally analyzed for identification of functional regions, and for the scanning of nucleotide sequence databases for errors in the sequences and the functional assignments associated with them. In an exhaustive scan of the GenBank database, 97.5% of the 744 known tRNA genes were correctly identified (true-positives), and 42 previously unidentified sequences were predicted to be tRNAs. A detailed analysis of these latter predictions reveals that 16 of the 42 are very similar to known tRNA genes, and we predict that they do, in fact, code for tRNA, yielding a false-positive rate for the algorithm of 0.003%. The new algorithm and testing strategy are a considerable improvement over any previously described strategies for recognizing tRNA genes, and they allow detections of genes (including introns) embedded in long genomic sequences.
Draft genome sequence of strain ATCC 35874 isolated from infected red oak in Washington, DC.
Shao J, Guan W, Zhao T, Huang Q Microbiol Resour Announc. 2023; 13(1):e0089323.
PMID: 38038447 PMC: 10793319. DOI: 10.1128/MRA.00893-23.
Guan W, Shao J, Zhao T, Huang Q Microbiol Resour Announc. 2022; 12(1):e0083122.
PMID: 36448819 PMC: 9872675. DOI: 10.1128/mra.00831-22.
Phage Genome Annotation: Where to Begin and End.
Shen A, Millard A Phage (New Rochelle). 2022; 2(4):183-193.
PMID: 36159890 PMC: 9041514. DOI: 10.1089/phage.2021.0015.
tRNAscan-SE 2.0: improved detection and functional classification of transfer RNA genes.
Chan P, Lin B, Mak A, Lowe T Nucleic Acids Res. 2021; 49(16):9077-9096.
PMID: 34417604 PMC: 8450103. DOI: 10.1093/nar/gkab688.
Conserved long-range base pairings are associated with pre-mRNA processing of human genes.
Kalmykova S, Kalinina M, Denisov S, Mironov A, Skvortsov D, Guigo R Nat Commun. 2021; 12(1):2300.
PMID: 33863890 PMC: 8052449. DOI: 10.1038/s41467-021-22549-7.