Search PubMed⌕ Search

PubMed · 14983071

Multiple structural alignment for distantly related all beta structures using TOPS pattern discovery and simulated annealing.

Abstract

Topsalign is a method that will structurally align diverse protein structures, for example, structural alignment of protein superfolds. All proteins within a superfold share the same fold but often have very low sequence identity and different biological and biochemical functions. There is often significant structural diversity around the common scaffold of secondary structure elements of the fold. Topsalign uses topological descriptions of proteins. A pattern discovery algorithm identifies equivalent secondary structure elements between a set of proteins and these are used to produce an initial multiple structure alignment. Simulated annealing is used to optimize the alignment. The output of Topsalign is a multiple structure-based sequence alignment and a 3D superposition of the structures. This method has been tested on three superfolds: the beta jelly roll, TIM (alpha/beta) barrel and the OB fold. Topsalign outperforms established methods on very diverse structures. Despite the pattern discovery working only on beta strand secondary structure elements, Topsalign is shown to align TIM (alpha/beta) barrel superfamilies, which contain both alpha helices and beta strands.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

A Williams, D R Gilbert, D R Westhead. 2003. Multiple structural alignment for distantly related all beta structures using TOPS pattern discovery and simulated annealing.. https://doi.org/10.1093/protein%2Fgzg116

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Packing helices in proteins by global optimization of a potential energy function.

An efficient method has been developed for packing alpha-helices in proteins. It treats alpha-helices as rigid bodies and uses a simplified Lennard-Jones potential with Miyazawa-Jernigan contact-energy parameters to describe the interactions between the alpha-helical elements in this coarse-grained system. Global conformational searches to generate packing arrangements rapidly are carried out with a Monte Carlo-with-minimization type of approach. The results for 42 proteins show that the approach reproduces native-like folds of alpha-helical proteins as low-energy local minima of this highly simplified potential function.

Protein Structure, Secondary↗

Structure and function of alpha-fetoprotein: a biophysical overview.

alpha-Fetoprotein (AFP) is a large serum glycoprotein belonging to the intriguing class of onco-developmental proteins. AFP has attracted considerable attention since it was shown that the change in its serum level during pregnancy is a hallmark of the development of numerous embryonic disorders, while the increase in its content in the plasma of adults correlates with the appearance of several pathological conditions. Over the past 30 years, some 11000 papers have been published concerning AFP, an average rate of over a publication a day since 1969. The majority of publications are about the application of the protein in diagnostics, or about other uses of AFP in biomedicine; though some of them describe the biochemical and functional properties of AFP, two aspects have been extensively reviewed. However, surprisingly little is currently known about structural properties of this protein as well as about the molecular mechanism of its function. The present review pursues the aim to describe the current state of the art in studies of structural properties and conformational stability of AFP. An attempt to establish the relationship between conformational transformations in AFP and its function is also made.

Protein Structure, Secondary↗

Closed loops of nearly standard size: common basic element of protein structure.

By screening the crystal protein structure database for close Calpha-Calpha contacts, a size distribution of the closed loops is generated. The distribution reveals a maximum at 27+/-5 residues, the same for eukaryotic and prokaryotic proteins. This is apparently a consequence of polymer statistic properties of protein chain trajectory. That is, closure into the loops depends on the flexibility (persistence length) of the chain. The observed preferential loop size is consistent with the theoretical optimal loop closure size. The mapping of the detected unit-size loops on the sequences of major typical folds reveals an almost regular compact consecutive arrangement of the loops. Thus, a novel basic element of protein architecture is discovered; structurally diverse closed loops of the particular size.

Protein Structure, Secondary↗