Search PubMedSearch

PubMed · 42547839

The shape of fitness functions and the distribution of mutational effect sizes jointly limit adaptation by regulatory mutations.

Abstract

Mutations in gene regulatory regions have been shown to play a role in rapid adaptation, but the factors determining their contribution are largely unknown. Here, using the metabolic enzyme cytosine deaminase of budding yeast, we examine whether adaptation to 5-fluorocytosine, which requires reduced cytosine deamination and can readily arise from amino acid substitutions, may be reached by single promoter mutations. We generated all single-nucleotide substitutions and indels in the FCY1 promoter and assayed the resulting mutants in presence of 5-fluorocytosine. This revealed that no promoter mutation is sufficient for adaptation to occur. We next investigated how this inaccessibility of adaptation arises by combining large-scale expression measurements with the experimental characterization of the corresponding expression-fitness function. These experiments showed that the shape of this function precludes single promoter mutations from being adaptive. Although 24% of mutations significantly affect expression, the fitness curve is flat around wild-type level. As such, adaptation can only emerge from a severe reduction of expression, which cannot occur from a single mutation in the promoter. Our results show that the contribution of regulatory mutations to rapid adaptation depends not only on the distribution of mutational effect sizes on expression level but also on the shape of the function linking fitness to expression levels.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Simon Aubé, Alexandre K Dubé, Christian R Landry. 2026-08-03. The shape of fitness functions and the distribution of mutational effect sizes jointly limit adaptation by regulatory mutations.. https://doi.org/10.1038/s41559-026-03142-x

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Modular synthetic cross-kingdom promoters enable coordinated expression in Escherichia coli and Saccharomyces cerevisiae.

Synthetic biology and metabolic engineering increasingly demand predictable and interoperable gene expression across phylogenetically distant organisms, as the need for portable genetic systems and transferable metabolic pathways continues to grow. However, fundamental differences in promoter architecture and transcriptional logic across kingdoms remain a key bottleneck in developing universal expression platforms. Here, we designed a set of modular hybrid promoters that enable tunable and quantitatively consistent gene expression in both Escherichia coli and Saccharomyces cerevisiae. These promoters integrate bacterial -10/-35 motifs and Shine-Dalgarno sequences with minimal yeast TATA boxes and Kozak sequences to ensure transcriptional and translational compatibility. The promoter set supported weak, moderate, and strong expression with high relative consistency across species. Applied to the biosynthetic pathway for the valuable pigment prodeoxyviolacein, the hybrid promoters enabled coordinated production in both hosts. This work establishes a broadly compatible promoter architecture and provides a foundational toolkit for cross-kingdom, multi-host synthetic biology.

Promoter Regions, Genetic

Structural Features of DNA in TATA-Containing and TATA-Less Core Promoters of RNA Polymerase II Differ.

Nucleotide motifs in the core promoters of eukaryotic protein-coding genes transcribed by RNA polymerase II (Pol II) play an important role in the transcription process. We analyzed the role of an octanucleotide located in the TATA box position. Depending on whether this octanucleotide can form a complex with the TATA-binding protein (TBP), the promoter is classified as either TATA-containing or TATA-less. We analyzed the differences in the primary and spatial structures, as well as their dynamics, in TATA-containing and TATA-less promoters of mammals and plants. We divided the complete promoter sets of six organisms (H. sapiens, M. musculus, C. familiaris, A. thaliana, Z. mays, and H. vulgare) from the EPDnew database into TATA-containing and TATA-less fractions. The sizes of the TATA-containing promoter fractions are significantly smaller than those of the TATA-less fractions in all studied organisms, except in A. thaliana, where the sizes of both fractions are approximately equal. We characterized promoter architecture using variation profiles of various base-pair step parameters, minor-groove width, and the conformational dynamics of native DNA. The architectures of TATA-containing and TATA-less promoters differ significantly. The possible mechanistic influence of DNA structural features on the formation of the pre-initiation complex (PIC) in both types of promoters is discussed.

Promoter Regions, Genetic

EvoSNR-Prom: Predicting promoters at single-nucleotide resolution with label-aware transfer learning of the pretrained EVO model.

The precise identification of promoters is crucial for understanding gene regulation. Deep learning methods have achieved considerable success in promoter prediction, yet most operate at the sequence level with coarse-grained labels. This means they label an entire DNA segment as either a "promoter" or "non-promoter," which results in a lack of the nucleotide-level resolution in prediction. In this study, we propose EvoSNR-Prom, a model designed for promoter prediction at single-nucleotide resolution. EvoSNR-Prom is built on the Evo foundation model and formulates promoter identification as a token-level sequence labeling problem, analogous to named entity recognition in natural language processing. To address the limited contextual information available in single-nucleotide tokenization, we introduce a lexicon-enhanced embedding strategy that incorporates biologically meaningful DNA lexicons, enriching contextual representations and improving the model's ability to capture complex sequence motifs. Furthermore, to enhance predictive performance on small size datasets, we integrate a label-aware transfer learning framework to leverage knowledge from well-annotated source species to a target organism. The results across various prokaryotic datasets show that EvoSNR-Prom achieves excellent performance. This work provides a valuable computational framework for the high-precision analysis of gene regulatory elements, contributing to the advancement of promoter prediction at single-nucleotide resolution.

Promoter Regions, Genetic