Search PubMed⌕ Search

PubMed · 12589754

Annotating nucleic acid-binding function based on protein structure.

Abstract

Many of the targets of structural genomics will be proteins with little or no structural similarity to those currently in the database. Therefore, novel function prediction methods that do not rely on sequence or fold similarity to other known proteins are needed. We present an automated approach to predict nucleic-acid-binding (NA-binding) proteins, specifically DNA-binding proteins. The method is based on characterizing the structural and sequence properties of large, positively charged electrostatic patches on DNA-binding protein surfaces, which typically coincide with the DNA-binding-sites. Using an ensemble of features extracted from these electrostatic patches, we predict DNA-binding proteins with high accuracy. We show that our method does not rely on sequence or structure homology and is capable of predicting proteins of novel-binding motifs and protein structures solved in an unbound state. Our method can also distinguish NA-binding proteins from other proteins that have similar, large positive electrostatic patches on their surfaces, but that do not bind nucleic acids.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Eric W Stawiski, Lydia M Gregoret, Yael Mandel-Gutfreund. 2003-02-28. Annotating nucleic acid-binding function based on protein structure.. https://doi.org/10.1016/s0022-2836(03)00031-7

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

GAMMA: gap-aware motif mining under incomplete labeling with applications to MHC motifs.

MOTIVATION: Sequence motif identification is crucial for understanding molecular recognition, particularly in immune responses involving peptide binding to major histocompatibility complex (MHC) Class I molecules for antigen presentation to T cells. Traditionally, MHC Class I binding motifs are assumed to be contiguous and span nine amino acids. However, structural evidence suggests that binding may involve nonadjacent residues, challenging the assumptions of existing methods. RESULTS: In this study, we propose Gap-Aware Motif Mining Algorithm (GAMMA), a probabilistic framework designed to identify noncontiguous motifs under conditions of incomplete labeling. GAMMA employs Bayesian inference with Markov chain Monte Carlo sampling to jointly estimate motif parameters, binding locations, and the relative spacing between binding positions. Through extensive simulations and real-world applications to MHC Class I peptide datasets, GAMMA outperforms existing motif discovery tools such as GLAM2 in accurately localizing binding residues and identifying the underlying motifs. Notably, our results suggest that the true number of binding residues may be eight, fewer than the commonly assumed nine. In addition, for longer peptides, the model captures increased flexibility in the central region, consistent with structural observations that peptides may bulge in the middle. AVAILABILITY AND IMPLEMENTATION: The raw data and the source codes are available on GitHub (https://github.com/RanLIUaca/GAMMAmotif).

Amino Acid Motifs↗

Signalling thresholds and negative B-cell selection in acute lymphoblastic leukaemia.

B cells are selected for an intermediate level of B-cell antigen receptor (BCR) signalling strength: attenuation below minimum (for example, non-functional BCR) or hyperactivation above maximum (for example, self-reactive BCR) thresholds of signalling strength causes negative selection. In ∼25% of cases, acute lymphoblastic leukaemia (ALL) cells carry the oncogenic BCR-ABL1 tyrosine kinase (Philadelphia chromosome positive), which mimics constitutively active pre-BCR signalling. Current therapeutic approaches are largely focused on the development of more potent tyrosine kinase inhibitors to suppress oncogenic signalling below a minimum threshold for survival. We tested the hypothesis that targeted hyperactivation--above a maximum threshold--will engage a deletional checkpoint for removal of self-reactive B cells and selectively kill ALL cells. Here we find, by testing various components of proximal pre-BCR signalling in mouse BCR-ABL1 cells, that an incremental increase of Syk tyrosine kinase activity was required and sufficient to induce cell death. Hyperactive Syk was functionally equivalent to acute activation of a self-reactive BCR on ALL cells. Despite oncogenic transformation, this basic mechanism of negative selection was still functional in ALL cells. Unlike normal pre-B cells, patient-derived ALL cells express the inhibitory receptors PECAM1, CD300A and LAIR1 at high levels. Genetic studies revealed that Pecam1, Cd300a and Lair1 are critical to calibrate oncogenic signalling strength through recruitment of the inhibitory phosphatases Ptpn6 (ref. 7) and Inpp5d (ref. 8). Using a novel small-molecule inhibitor of INPP5D (also known as SHIP1), we demonstrated that pharmacological hyperactivation of SYK and engagement of negative B-cell selection represents a promising new strategy to overcome drug resistance in human ALL.

Amino Acid Motifs↗

Purification and cDNA cloning of a novel antibacterial peptide with a cysteine-stabilized alphabeta motif from the longicorn beetle, Acalolepta luxuriosa.

An antibacterial peptide from the hemolymph of a coleopteran insect, Acalolepta luxuriosa, in the superfamily Cerambyocidea was characterized. The mature antibacterial peptide had 27 amino acid residues with a theoretical molecular weight of 3099.29 and it showed antibacterial activity against Escherichia coli and Micrococcus luteus. The deduced amino acid sequence of the peptide showed that it had a cysteine-stabilized alphabeta motif with a C...CXXXC...C...CXC consensus sequence, like insect defensins. However, the results of a multiple sequence alignment and phylogenetic analysis with CLUSTAL X indicated that this peptide is a novel peptide with a cysteine-stabilized alphabeta motif that is distant from insect defensins.

Amino Acid Motifs↗