PubMed · 9415986
Computational gene identification: an open problem.
Abstract
As the Human Genome Project enters the large-scale sequencing phase, computational gene identification methods are becoming essential for the automatic analysis and annotation of large uncharacterized genomic sequences. Currently available computer programs relying mainly on sequence coding statistics are of great use in pin-pointing regions in genomic sequences containing exons. Such programs perform rather poorly, however, when the problem is to fully elucidate gene structure. For this problem, the DNA sequence signals involved in the specification of the genes--start sites and splice sites--carry a lot of information, and simple methods relying on such information can predict gene structure with an accuracy to some extent comparable to that of other more sophisticated computational methods.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
R Guigó. 1997. Computational gene identification: an open problem.. https://doi.org/10.1016/s0097-8485(97)00008-9
Cite the original work for its findings. Save a collection to share your selection of sources.