Search PubMed⌕ Search

PubMed · 14649297

Structure-based functional inference in structural genomics.

Abstract

The dramatically increasing number of new protein sequences arising from genomics and proteomics requires the need for methods to rapidly and reliably infer the molecular and cellular functions of these proteins. One such approach, structural genomics, aims to delineate the total repertoire of protein folds in nature, thereby providing three-dimensional folding patterns for all proteins and to infer molecular functions of the proteins based on the combined information of structures and sequences. The goal of obtaining protein structures on a genomic scale has motivated the development of high throughput technologies and protocols for macromolecular structure determination that have begun to produce structures at a greater rate than previously possible. These new structures have revealed many unexpected functional inferences and evolutionary relationships that were hidden at the sequence level. Here, we present samples of structures determined at Berkeley Structural Genomics Center and collaborators' laboratories to illustrate how structural information provides and complements sequence information to deduce the functional inferences of proteins with unknown molecular functions. Two of the major premises of structural genomics are to discover a complete repertoire of protein folds in nature and to find molecular functions of the proteins whose functions are not predicted from sequence comparison alone. To achieve these objectives on a genomic scale, new methods, protocols, and technologies need to be developed by multi-institutional collaborations worldwide. As part of this effort, the Protein Structure Initiative has been launched in the United States (PSI; www.nigms.nih.gov/funding/psi.html). Although infrastructure building and technology development are still the main focus of structural genomics programs, a considerable number of protein structures have already been produced, some of them coming directly out of semiautomated structure determination pipelines. The Berkeley Structural Genomics Center (BSGC) has focused on the proteins of Mycoplasma or their homologues from other organisms as its structural genomics targets because of the minimal genome size of the Mycoplasmas as well as their relevance to human and animal pathogenicity (http://www.strgen.org). Here we present several protein examples encompassing a spectrum of functional inferences obtainable from their three-dimensional structures in five situations, where the inferences are new and testable, and are not predictable from protein sequence information alone.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Sung-Hou Kim, Dong Hae Shin, In-Geol Choi, Ursula Schulze-Gahmen, Shengfeng Chen, Rosalind Kim. 2003. Structure-based functional inference in structural genomics.. https://doi.org/10.1023/a%3A1026200610644

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The human mitochondrial genome contains a second light strand promoter.

The human mitochondrial genome must be replicated and expressed in a timely manner to maintain energy metabolism and supply cells with adequate levels of adenosine triphosphate. Central to this process is the idea that replication primers and gene products both arise via transcription from a single light strand promoter (LSP) such that primer formation can influence gene expression, with no consensus as to how this is regulated. Here, we report the discovery of a second light strand promoter (LSP2) in humans, with features characteristic of a bona fide mitochondrial promoter. We propose that the position of LSP2 on the mitochondrial genome allows replication and gene expression to be orchestrated from two distinct sites, which expands our long-held understanding of mitochondrial gene expression in humans.

Adenosine Triphosphate↗

GCN2 kinase activation by ATP-competitive kinase inhibitors.

Small-molecule kinase inhibitors represent a major group of cancer therapeutics, but tumor responses are often incomplete. To identify pathways that modulate kinase inhibitor response, we conducted a genome-wide knockout (KO) screen in glioblastoma cells treated with the pan-ErbB inhibitor neratinib. Loss of general control nonderepressible 2 (GCN2) kinase rendered cells resistant to neratinib, whereas depletion of the GADD34 phosphatase increased neratinib sensitivity. Loss of GCN2 conferred neratinib resistance by preventing binding and activation of GCN2 by neratinib. Several other Food and Drug Administration (FDA)-approved inhibitors, such erlotinib and sunitinib, also bound and activated GCN2. Our results highlight the utility of genome-wide functional screens to uncover novel mechanisms of drug action and document the role of the integrated stress response (ISR) in modulating the response to inhibitors of oncogenic kinases.

Adenosine Triphosphate↗

Studies of SpoIIAB mutant proteins elucidate the mechanisms that regulate the developmental transcription factor sigmaF in Bacillus subtilis.

SigmaF, the first compartment-specific sigma factor of sporulation, is regulated by an anti-sigma factor, SpoIIAB (AB) and its antagonist SpoIIAA (AA). AB can bind to sigmaF in the presence of ATP or to AA in the presence of ADP; in addition, AB can phosphorylate AA. The ability of AB to switch between its two binding partners regulates sigmaF. Early in sporulation, AA activates sigmaF by releasing it from its complex with AB. We have previously proposed a reaction scheme for the phosphorylation of AA by AB which accounts for AA's regulatory role. A crucial feature of this scheme is a conformational change in AB that accompanies its switch in binding partner. In the present study, we have studied three AB mutants, all of which have amino-acid replacements in the nucleotide-binding region; AB-E104K (Glu104-->Lys) and AB-T49K (Thr49-->Lys) fail to activate sigmaF, and AB-R105A (Arg105-->Ala) activates it prematurely. We used techniques of enzymology, surface plasmon resonance and fluorescence spectroscopy to analyse the defects in each mutant. AB-E104K was deficient in binding to AA, AB-T49K was deficient in binding to ADP and AB-R105A bound ADP exceptionally strongly. Although the release of sigmaF from all three mutant proteins was impaired, and all three failed to undergo the wild-type conformational change when switching binding partners, the phenotypes of the mutant cells were best accounted for by the properties of the respective AB species in forming complexes with AA and ADP. The behaviour of the mutants enables us to propose convincing mechanisms for the regulation of sigmaF in wild-type bacteria.

Adenosine Triphosphate↗