Search PubMed⌕ Search

PubMed · 10739263

Expectations from structural genomics.

Abstract

Structural genomics projects aim to provide an experimental structure or a good model for every protein in all completed genomes. Most of the experimental work for these projects will be directed toward proteins whose fold cannot be readily recognized by simple sequence comparison with proteins of known structure. Based on the history of proteins classified in the SCOP structure database, we expect that only about a quarter of the early structural genomics targets will have a new fold. Among the remaining ones, about half are likely to be evolutionarily related to proteins of known structure, even though the homology could not be readily detected by sequence analysis.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

S E Brenner, M Levitt. 2000. Expectations from structural genomics.. https://doi.org/10.1110/ps.9.1.197

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Protein structure comparison using the markov transition model of evolution.

A number of automatic protein structure comparison methods have been proposed; however, their similarity score functions are often decided by the researchers' intuition and trial-and-error, and not by theoretical background. We propose a novel theory to evaluate protein structure similarity, which is based on the Markov transition model of evolution. Our similarity score between structures i and j is defined as log P(j --> i)/P(i), where P(j --> i) is the probability that structure j changes to structure i during the evolutionary process, and P(i) is the probability that structure i appears by chance. This is a reasonable definition of structure similarity, especially for finding evolutionarily related (homologous) similarity. The probability P(j --> i) is estimated by the Markov transition model, which is similar to the Dayhoff's substitution model between amino acids. To estimate the parameters of the model, homologous protein structure pairs are collected using sequence similarity, and the numbers of structure transitions within the pairs are counted. Next these numbers are transformed to a transition probability matrix of the Markov transition. Transition probabilities for longer time are obtained by multiplying the probability matrix by itself several times. In this study, we generated three types of structure similarity scores: an environment score, a residue-residue distance score, and a secondary structure elements (SSE) score. Using these scores, we developed the structure comparison program, Matras (MArkovian TRAnsition of protein Structure). It employs a hierarchical alignment algorithm, in which a rough alignment is first obtained by SSEs, and then is improved with more detailed functions. We attempted an all-versus-all comparison of the SCOP database, and evaluated its ability to recognize a superfamily relationship, which was manually assigned to be homologous in the SCOP database. A comparison with the FSSP database shows that our program can recognize more homologous similarity than FSSP. We also discuss the reliability of our method, by studying the disagreement between structural classifications by Matras and SCOP.

Databases, Factual↗

Identification of lorazepam and sildenafil as examples for the application of LC/ionspray-MS and MS-MS with mass spectra library searching in forensic toxicology.

A mass spectra (MS) library using in-source collision induced dissociation (ESI-CID) as well as a tandem-mass spectra (MS-MS) library with product ion spectra of drugs has recently been developed with a triple-quadrupole ionspray mass spectrometer [1,2]. For the ESI-CID MS library, single-quadrupole mode and for the MS-MS library triple-quadrupole mode have been used. These mass spectra libraries were applied successfully for the general-unknown screening for drugs and metabolites in serum and urine with liquid-chromatography-mass spectrometry (LC-MS) using a PE/SCIEX API 365 with a turboionspray source. As examples, the identification of lorazepam and lorazepam-glucuronide in a serum extract and the identification of sildenafil and alkyloxidated sildenafil in urine are presented here.

Databases, Factual↗