Search PubMed⌕ Search

PubMed · 16019027

Protein function prediction using local 3D templates.

Abstract

The prediction of a protein's function from its 3D structure is becoming more and more important as the worldwide structural genomics initiatives gather pace and continue to solve 3D structures, many of which are of proteins of unknown function. Here, we present a methodology for predicting function from structure that shows great promise. It is based on 3D templates that are defined as specific 3D conformations of small numbers of residues. We use four types of template, covering enzyme active sites, ligand-binding residues, DNA-binding residues and reverse templates. The latter are templates generated from the target structure itself and scanned against a representative subset of all known protein structures. Together, the templates provide a fairly thorough coverage of the known structures and ensure that if there is a match to a known structure it is unlikely to be missed. A new scoring scheme provides a highly sensitive means of discriminating between true positive and false positive template matches. In all, the methodology provides a powerful new tool for function prediction to complement those already in use.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Roman A Laskowski, James D Watson, Janet M Thornton. 2005-08-19. Protein function prediction using local 3D templates.. https://doi.org/10.1016/j.jmb.2005.05.067

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Normal modes of redox-active tyrosine: conformation dependence and comparison to experiment.

Redox-active tyrosine residues play important roles in long-distance electron reactions in enzymes such as prostaglandin H synthase, ribonucleotide reductase, and photosystem II (PSII). Spectroscopic characterization of tyrosyl radicals in these systems provides a powerful experimental probe into the role of the enzyme in mediation of long-range electron transfer processes. Interpretation of such data, however, relies critically on first establishing a spectroscopic fingerprint of isotopically labeled tyrosinate and tyrosyl radicals in nonenzymatic environments. In this report, FT-IR results obtained from tyrosinate, tyrosyl radical (produced by ultraviolet photolysis of polycrystalline tyrosinate), and their isotopologues at 77 K are presented. Assignment of peaks and isotope shifts is aided by density-functional B3LYP/6-311++G(3df,2p)//B3LYP/6-31++G(d,p) calculations of tyrosine and tyrosyl radical in several different charge and protonation states. In addition, characterization of the potential energy surfaces of tyrosinate and tyrosyl radical as a function of the backbone and ring torsion angles provides detailed insight into the sensitivity of the vibrational frequencies to conformational changes. These results provide a detailed spectroscopic interpretation, which will elucidate the structures of redox-active tyrosine residues in complex protein environments. Specific application of these data is made to enzymatic systems.

Models, Molecular↗

Synthesis and biological evaluation of type VI beta-turn templated RGD peptidomimetics.

We report the design, synthesis, and binding affinities of a family of cyclic RGD peptides attached to type VI beta-turn scaffolds. The analogues prepared exhibit interesting binding data to the isolated receptors alphavbeta3 and alphavbeta5. The results demonstrate the utility of these type VI beta-turn scaffolds for the constraint of biologically relevant peptides.

Models, Molecular↗

A composite score for predicting errors in protein structure models.

Reliable prediction of model accuracy is an important unsolved problem in protein structure modeling. To address this problem, we studied 24 individual assessment scores, including physics-based energy functions, statistical potentials, and machine learning-based scoring functions. Individual scores were also used to construct approximately 85,000 composite scoring functions using support vector machine (SVM) regression. The scores were tested for their abilities to identify the most native-like models from a set of 6000 comparative models of 20 representative protein structures. Each of the 20 targets was modeled using a template of <30% sequence identity, corresponding to challenging comparative modeling cases. The best SVM score outperformed all individual scores by decreasing the average RMSD difference between the model identified as the best of the set and the model with the lowest RMSD (DeltaRMSD) from 0.63 A to 0.45 A, while having a higher Pearson correlation coefficient to RMSD (r=0.87) than any other tested score. The most accurate score is based on a combination of the DOPE non-hydrogen atom statistical potential; surface, contact, and combined statistical potentials from MODPIPE; and two PSIPRED/DSSP scores. It was implemented in the SVMod program, which can now be applied to select the final model in various modeling problems, including fold assignment, target-template alignment, and loop modeling.

Models, Molecular↗