Search PubMed⌕ Search

Biomedical subjects

Takuyo Aita

Publications and source records attributed to Takuyo Aita.

10 recordsLinked to original sources

Experimental rugged fitness landscape in protein sequence space.

The fitness landscape in sequence space determines the process of biomolecular evolution. To plot the fitness landscape of protein function, we carried out in vitro molecular evolution beginning with a defective fd phage carrying a random polypeptide of 139 amino acids in place of the g3p minor coat protein D2 domain, which is essential for phage infection. After 20 cycles of random substitution at sites 12-130 of the initial random polypeptide and selection for infectivity, the selected phage showed a 1.7x10(4)-fold increase in infectivity, defined as the number of infected cells per ml of phage suspension. Fitness was defined as the logarithm of infectivity, and we analyzed (1) the dependence of stationary fitness on library size, which increased gradually, and (2) the time course of changes in fitness in transitional phases, based on an original theory regarding the evolutionary dynamics in Kauffman's n-k fitness landscape model. In the landscape model, single mutations at single sites among n sites affect the contribution of k other sites to fitness. Based on the results of these analyses, k was estimated to be 18-24. According to the estimated parameters, the landscape was plotted as a smooth surface up to a relative fitness of 0.4 of the global peak, whereas the landscape had a highly rugged surface with many local peaks above this relative fitness value. Based on the landscapes of these two different surfaces, it appears possible for adaptive walks with only random substitutions to climb with relative ease up to the middle region of the fitness landscape from any primordial or random sequence, whereas an enormous range of sequence diversity is required to climb further up the rugged surface above the middle region.

Amino Acid Sequence↗

Directed evolution by accumulating tailored mutations: thermostabilization of lactate oxidase with less trade-off with catalytic activity.

We assumed that adverse effects posed by introducing multiple mutations could be decomposed into those of each of the component mutations and that the risk could be reduced by the accumulation of mutations that were finely tuned for directed improvement of a specific property. We propose here a directed evolution strategy for improving a specific property with less effect on other ones. This strategy is composed of fine-tuning of mutations and their accumulation by our original mutation-assembling method. In this study, we selected lactate oxidase (LOX) as a model enzyme, because its directed evolution had showed a trade-off between thermostability and catalytic activity. Mutation profiling at each of the sites found by error-prone PCR revealed a strong inverse relationship between the two properties. Thermostable mutations with less effect on catalytic activity were selected at each site and accumulated with ideal combinations by our method. The resultant multiple mutants exhibited 5- to 10-fold superior catalytic activity and comparable thermostability with those created by accumulating thermostable mutations, which were not tuned for catalytic activity. This result demonstrates that the accumulation of fine-tuned mutations is an advantageous approach to reduce the risk of adverse effects posed by accumulating multiple mutations.

Amino Acid Substitution↗

Modified substrate specificity of pyrroloquinoline quinone glucose dehydrogenase by biased mutation assembling with optimized amino acid substitution.

A biased mutation-assembling method-that is, a directed evolution strategy to facilitate an optimal accumulation of multiple mutations on the basis of additivity principles, was applied to the directed evolution of water-soluble PQQ glucose dehydrogenase (PQQGDH-B) to reduce its maltose oxidation activity, which can lead to errors in blood glucose determination. Mutations appropriate for the reduction without fatal deterioration of its glucose oxidation activity were developed by an error-prone PCR method coupled with a saturation mutagenesis method. Moreover, two types of incorporation frequency based on their contribution were assigned to the mutations: high (80%) and evens (50%), in constructing a multiple mutant library. The best mutant created showed a marked reduction in maltose oxidation activity, corresponding to 4% of that of the wild-type enzyme, with 35% retention of glucose oxidation activity. In addition, this mutant showed a reduction in galactose oxidation activity corresponding to 5% of that of the wild-type enzyme. In conclusion, we succeeded in developing the PQQGDH-B mutants with improved substrate specificity and validated our method coupled with optimized mutations and their contribution-based incorporation frequencies by applying it to the development.

Acinetobacter calcoaceticus↗

Biased mutation-assembling: an efficient method for rapid directed evolution through simultaneous mutation accumulation.

We have developed an efficient optimization technique, 'biased mutation-assembling', for improving protein properties such as thermostability. In this strategy, a mutant library is constructed using the overlap extension polymerase chain reaction technique with DNA fragments from wild-type and phenotypically advantageous mutant genes, in which the number of mutations assembled in the wild-type gene is stochastically controlled by the mixing ratio of the mutant DNA fragments to wild-type fragments. A high mixing ratio results in a mutant composition biased to favor multiple-point mutants. We applied this strategy to improve the thermostability of prolyl endopeptidase from Flavobacterium meningosepticum as a case study and found that the proportion of thermostable mutants in a library increased as the mixing ratio was increased. If the proportion of thermostable mutants increases, the screening effort needed to find them should be reduced. Indeed, we isolated a mutant with a 1200-fold longer activity half-life at 60 degrees C than that of wild-type prolyl endopeptidase after screening only 2000 mutants from a library prepared with a high mixing ratio. Our results indicate that an aggressive accumulation of advantageous mutations leads to an increase in the quality of the mutant library and a reduction in the screening effort required to find superior mutants.

Chryseobacterium↗

Thermodynamical interpretation of evolutionary dynamics on a fitness landscape in an evolution reactor, II.

In our previous report [Aita, T., Morinaga, S., Hosimi, Y., 2004. Thermodynamical interpretation of evolutionary dynamics on a fitness landscape in an evolution reactor I. Bull. Math. Biol. 66, 1371-1403], an analogy between thermodynamics and adaptive walks on a Mt. Fuji-type fitness landscape in an artificial selection system was presented. Introducing the 'free fitness' as the sum of a fitness term and an entropy term and 'evolutionary force' as the gradient of free fitness on a fitness coordinate, we demonstrated that the adaptive walk (=evolution) is driven by the evolutionary force in the direction in which free fitness increases. In this report, we examine the effect of various modifications of the original model on the properties of the adaptive walk. The modifications were as follows: first, mutation distance d was distributed obeying binomial distribution; second, the selection process obeyed the natural selection protocol; third, ruggedness was introduced to the landscape according to the NK model; fourth, a noise was included in the fitness measurement. The effect of each modification was described in the same theoretical framework as the original model by introducing 'effective' quantities such as the effective mutation distance or the effective screening size.

Algorithms↗

Thermodynamical interpretation of evolutionary dynamics on a fitness landscape in a evolution reactor, I.

A theory for describing evolution as adaptive walks by a finite population with M walkers (M > or = 1) on an anisotropic Mt. Fuji-type fitness landscape is presented, from a thermodynamical point of view. Introducing the 'free fitness' as the sum of a fitness term and an entropy term and 'evolutionary force' as the gradient of free fitness on a fitness coordinate, we demonstrate that the behavior of these theoretical walkers is almost consistent with the thermodynamical schemes. The major conclusions are as follows: (1) an adaptive walk (=evolution) is driven by an evolutionary force in the direction in which free fitness increases; (2) the expectation of the climbing rate obeys an equation analogous to the Einstein relation in Brownian motion; (3) the standard deviation of the climbing rate is a quantity analogous to the mean thermal energy of a particle, kT (x constant). In addition, on the interpretation that the walkers climb the landscape by absorbing 'fitness information' from the surroundings, we succeeded in quantifying the fitness information and formulating a macroscopic scheme from an informational point of view.

Adaptation, Biological↗

Thermodynamical interpretation of an adaptive walk on a Mt. Fuji-type fitness landscape: Einstein relation-like formula holds in a stochastic evolution.

We have theoretically studied the statistical properties of adaptive walks (or hill-climbing) on a Mt. Fuji-type fitness landscape in the multi-dimensional sequence space through mathematical analysis and computer simulation. The adaptive walk is characterized by the "mutation distance" d as the step-width of the walker and the "population size" N as the number of randomly generated d-fold point mutants to be screened. In addition to the fitness W, we introduced the following quantities analogous to thermodynamical concepts: "free fitness" G(W) is identical with W+T x S(W), where T is the "evolutionary temperature" T infinity square root of d/lnN and S(W) is the entropy as a function of W, and the "evolutionary force" X is identical with d(G(W)/T)/dW, that is caused by the mutation and selection pressure. It is known that a single adaptive walker rapidly climbs on the fitness landscape up to the stationary state where a "mutation-selection-random drift balance" is kept. In our interpretation, the walker tends to the maximal free fitness state, driven by the evolutionary force X. Our major findings are as follows: First, near the stationary point W*, the "climbing rate" J as the expected fitness change per generation is described by J approximately L x X with L approximately V/2, where V is the variance of fitness distribution on a local landscape. This simple relationship is analogous to the well-known Einstein relation in Brownian motion. Second, the "biological information gain" (DeltaG/T) through adaptive walk can be described by combining the Shannon's information gain (DeltaS) and the "fitness information gain" (DeltaW/T).

Animals↗

An in silico exploration of the neutral network in protein sequence space.

Designating amino-acid sequences that fold into a common main-chain structure as "neutral sequences" for the structure, regardless of their function or stability, we investigated the distribution of neutral sequences in protein sequence space. For four distinct target structures (alpha, beta,alpha/beta and alpha+beta types) with the same chain length of 108, we generated the respective neutral sequences by using the inverse folding technique with a knowledge-based potential function. We assumed that neutral sequences for a protein structure have Z scores higher than or equal to fixed thresholds, where thresholds are defined as the Z score for the corresponding native sequence (case 1) or much greater Z score (case 2). An exploring walk simulation suggested that the neutral sequences mapped into the sequence space were connected with each other through straight neutral paths and formed an inherent neutral network over the sequence space. Through another exploring walk simulation, we investigated contiguous regions between or among the neutral networks for the distinct protein structures and obtained the following results. The closest approach distance between the two neutral networks ranged from 5 to 29 on the Hamming distance scale, showing a linear increase against the threshold values. The sequences located at the "interchange" regions between the two neutral networks have intermediate sequence-profile-scores for both corresponding structures. Introducing a "ball" in the sequence space that contains at least one neutral sequence for each of the four structures, we found that the minimal radius of the ball that is centered at an arbitrary position ranged from 35 to 50, while the minimal radius of the ball that is centered at a certain special position ranged from 20 to 30, in the Hamming distance scale. The relatively small Hamming distances (5-30) may support an evolution mechanism by transferring from a network for a structure to another network for a more beneficial structure via the interchange regions.

Amino Acid Sequence↗

Statistical formulae of the energy distribution among a globular protein structure ensemble.

In prediction of a protein main-chain structure into which a query sequence of amino acids folds, one evaluates the relative stability of a candidate structure against reference structures. We developed a statistical theory for calculating the energy distribution over a main-chain structure ensemble, only with an amino acid composition given as a single argument. Then, we obtained a statistical formulae of the ensemble mean and ensemble variance V[E] of the reference structural energies, as explicit functions of the amino acid composition. The mean and the variance V[E] calculated from the formulae were well or roughly consistent with those resulting from a gapless threading simulation. We can use the formulae not only to perform the high-through-put screening of sequences in the inverse folding problem, but also to handle the problem analytically.

Animals↗

Surveying a local fitness landscape of a protein with epistatic sites for the study of directed evolution.

We present a method for analysis of a fitness landscape of a biopolymer with significantly epistatic sites. The analysis is based on a quasi-additive fitness model. The fitness model is constructed with additive terms conducted by "site-fitness" and epistatic terms conducted by "pair-fitness," where the site-fitness is a fitness contribution from an independent residue and the pair-fitness is a fitness contribution from a pair of epistatic residues. As a case study, we analyzed the sequence-fitness data for 45 clones of thermostable prolyl endopeptidase mutants. They were generated by a mutation scrambling method, which can accumulate advantageous mutations. The fitness contributions from 14 single-point mutations including E67Q and Q656R were identified by the analysis. As a result, we found that the fitness model with a significant epistatic term by a pair of the 67th site and 656th site was in good agreement with the experimental data and that the explored landscape in the binary 14-dimensional sequence space is still a mountainous landscape with twin peaks. The validity was supported by the analysis of mutant fitness distributions derived from another mutation scrambling experiment and by (3D) structural data.

Directed Molecular Evolution↗