Search PubMedSearch

SEARCH · Search PubMed

Results for “Ploidy variation and whole-genome doubling”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

2,996 records · Page 22Linked to original sources

Ramu stunt virus genome reveals previously unreported segments and nucleocapsid domain duplication in Mechlorovirus.

Ramu stunt virus (RmSV), a member of the genus Mechlorovirus within the family Phenuiviridae, was previously described as a six-segmented RNA virus infecting sugarcane. In this study, we re-examined type material and additional isolates using high-throughput sequencing and RT-PCR validation, revealing that RmSV possesses a nine-segmented genome, making it the largest reported in the Phenuiviridae. This expanded architecture includes duplicated RNA segments (RNA 2a and RNA 2b) encoding nucleocapsid-like proteins and two novel segments (RNA 7 and RNA 8). Comparative analysis showed that RNA 2a and 2b share about 84% amino acid identity, while RNA 5 encodes a third nucleocapsid homolog, indicating unprecedented domain redundancy. Structural modeling confirmed that all three nucleocapsid proteins maintain a conserved fold despite low sequence identity, with electrostatic mapping suggesting differential RNA-binding potential. Additionally, RNA 6 encodes a hypothetical protein structurally similar to the rice stripe virus disease-specific S-protein, implicating a role in symptom development. Transcript abundance analysis revealed RNA 6 as the most highly expressed segment across isolates. These findings revise the genomic composition of RmSV, highlight mechanisms of genome plasticity and adaptive evolution in plant-infecting bunyaviruses, and underscore practical implications for diagnostic assay design, resistance breeding, and biosecurity surveillance.

Genome, Viral

Genomic and Phenotypic Characterization of Two Novel Enterobacter Phages With EDTA-Enhanced Antibiofilm Activity.

Multidrug-resistant members of the Enterobacter cloacae complex (ECC) are increasingly linked to difficult-to-treat infections and biofilm-mediated antimicrobial tolerance. Here, two lytic phages, vB_EhoIP_HHH and vB_EluM_RZH, displaying podovirus-like and myovirus-like morphology, respectively, were isolated from the River Chelt. HHH has a 39,582 bp genome (51.2% GC, 63 ORFs), while RZH has a 174,197 bp genome (39.4% GC, 314 ORFs), with neither genome carrying antimicrobial resistance, virulence or lysogeny-associated genes. VIRIDIC and VICTOR analyses placed HHH within Kayfunavirus and RZH within Karamvirus, supporting their classification as distinct species. Both phages demonstrated rapid adsorption, short latent periods and stability across physiological pH and temperature ranges. A phage cocktail targeting MDR ECC strain was evaluated with EDTA against established biofilms. Crystal violet assays showed the greatest biomass reduction at MOI 10 with 0.5-0.75 mM EDTA. Bliss independence analysis revealed localized synergy within this window but significant overall antagonism at higher EDTA concentrations. CFU enumeration confirmed greater activity against 24 h than 48 h biofilms. The optimized combination also reduced recoverable bacteria in a fibroblast infection model while maintaining low LDH release. These findings identify two novel lytic Enterobacter phages and support a narrow EDTA concentration window for enhanced phage-mediated antibiofilm activity.

Biofilms

Cost-Effectiveness and the Economics of Genomic Testing and Molecularly Matched Therapies.

Cost-effectiveness analysis of precision oncology can help guide value-driven care. Next-generation sequencing is increasingly cost-efficient over single gene testing because diagnostic algorithms require multiple individual gene tests to determine biomarker status. Matched targeted therapy is often not cost-effective due to the high cost associated with drug treatment. However, genomic profiling can promote cost-effective care by identifying patients who are unlikely to benefit from therapy. Additional applications of genomic profiling such as universal testing for hereditary cancer syndromes and germline testing in patients with cancer may represent cost-effective approaches compared with traditional history-based diagnostic methods.

Humans

Revealing the Shared Genetic Architecture of Metabolic Dysfunction-Associated Steatotic Liver Disease-Related Traits Through Genomic Structural Equation Modeling.

Although individual traits related to metabolic dysfunction-associated steatotic liver disease (MASLD) have been investigated through large-scale genome-wide association studies (GWASs), the shared genetic susceptibility across these traits remains unclear. We therefore conducted a multivariate GWAS of key MASLD-related traits to elucidate their common genetic architecture. We applied genomic structural equation modeling to model a latent genetic factor (MASLD-F) underlying genetically correlated MASLD-related traits, leveraging their GWAS-derived genetic correlations. We then performed functional annotations, including fine-mapping, transcriptome-wide association study, and cell- and tissue-type-specific enrichment analyses, and conducted Mendelian randomization analyses to identify modifiable risk factors. Our multivariate MASLD-F GWAS identified 50 independent variants across 48 genomic loci. Transcriptomic imputation identified several MASLD-F-associated genes, including ARNTL, NPC1, BTBD10, VDAC2, TSKU, SFMBT1, and ABHD17C. We observed significant enrichment of MASLD-F-related genetic signals predominantly in brain tissues, pancreatic islets, and the adrenal gland. Additionally, six modifiable risk factors and four modifiable protective factors for MASLD-F were identified. These findings reveal a complex shared genetic architecture underlying MASLD components, thereby expanding our understanding of disease pathogenesis and providing novel insights for precision medicine and public health interventions.

Humans

Complete genome sequence of multidrug-resistant Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from chicken liver, harboring a novel imipenem resistance mechanism.

We present the complete genome sequence of Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from Gallus gallus liver in China, harboring plasmid pSE191. The genome reveals multiple antibiotic resistance mechanisms and phenotypic imipenem resistance without canonical genes.

antibiotic resistance

Peptide molecular lock-engineered nanobodies enable an oriented dual-modal immunoassay for reliable detection of Cronobacter sakazakii.

Conventional nanobody ELISAs for trace Cronobacter sakazakii in powdered infant formula suffer from random orientation and low signal output. We developed an oriented dual-modal immunoassay that combines site-specific biotinylation via a C-terminal AviTag and a peptide molecular lock, enabling controlled surface orientation while preserving nanobody structural integrity. This strategy was further integrated with phage-displayed nanobodies for multivalent amplification and both fluorescent and colorimetric readouts. The assay exhibited a broad linear range of 103-106 CFU/mL, with limits of detection (LODs) of 6.70 × 102 CFU/mL for fluorescence and 1.55 × 103 CFU/mL for colorimetry, showing improved sensitivity compared with the conventional passive adsorption-based Nb-ELISA evaluated in this study. XGBoost-based multimodal fusion improved quantitative accuracy, and SHAP analysis elucidated modality contributions. In spiked powdered infant formula samples, recoveries ranged from 92.1% to 118% with coefficients of variation below 5.98%, confirming acceptable matrix tolerance and analytical reliability.

Cronobacter sakazakii

Introgression shapes the genomic conflict landscape of Malus, providing evidence for a reticulate backbone in a woody crop lineage.

Phylogenomic discordance is widespread across plants, but its evolutionary significance is often obscured when conflict is treated primarily as analytical noise rather than as evidence of underlying processes. In woody lineages in particular, incomplete lineage sorting, introgression, and genome duplication can interact over long timescales to produce complex genomic histories that are not adequately summarized by a strictly bifurcating tree. Here, we use Malus as a model woody genus to investigate how these processes structure conflict across a genus-scale, accession-based phylogenomic framework. Using broad taxon sampling, hundreds of nuclear loci, plastid genomes, and genome-wide SNP summaries, we reconstruct a robust nuclear backbone for sampled Malus lineages and evaluate where discordance is concentrated and which processes best explain it. Nuclear analyses resolve eight major clades, whereas conflict is non-random and localized to recurrent hotspots rather than evenly distributed across the tree. Cytonuclear discordance is similarly concentrated, especially around Clade H, represented by sampled accessions of M. tschonoskii, where localized plastid-nuclear disagreement is consistent with candidate plastid capture or organellar introgression. Multiple complementary analyses further indicate that the strongest conflict is not explained by ILS alone, but instead reflects lineage-structured introgression, while polyploid complexes represent additional localized sources of evolutionary complexity. Together, these results provide evidence for a reticulate genomic backbone in Malus and show how integrating nuclear, plastid, and genome-wide conflict analyses can help distinguish background discordance from process-specific signals in woody plant radiations. Several lineage-level reticulation hypotheses identified here should now be tested with broader population-level sampling and curated reference accessions.

Malus

Shared genetic architecture and cellular convergence between female reproductive disorders and pulmonary function: a genome-wide cross-trait analysis.

Female reproductive disorders (FRDs), including polycystic ovary syndrome, endometriosis, uterine leiomyomata, and infertility, have been epidemiologically associated with impaired pulmonary function. However, it remains unclear whether this cross-organ link reflects shared genetic etiology and, if so, which cellular mechanisms mediate it. We performed a systematic genome-wide cross-trait analysis of three FRDs and lung function traits (FEV₁, FVC, FEV₁/FVC) using GWAS summary statistics from individuals of European ancestry, integrating genetic correlation, bidirectional causal inference, pleiotropy mapping, and single-cell enrichment analyses. We identified significant negative genetic correlations between FRDs and lung volume traits, most prominently for FVC (rg range: - 0.077 to - 0.178). Bidirectional causal analyses indicated that FRDs have a detrimental effect on lung volume, with higher FRD genetic liability associated with reduced lung volume. Cross-trait meta-analysis identified 17 pleiotropic variants across 11 loci, with the 19q13.2 (LTBP4) and 12q13.13 (HOXC6/HOXC9) loci showing strong evidence of shared causal variants. Critically, single-cell analyses revealed that shared genetic risk converged on mesenchymal lineages across organs, specifically alveolar adventitial fibroblasts in the lung and stromal/smooth muscle cells in the endometrium. Transcriptome-wide analyses further nominated the estrogen-responsive gene RERG as a convergent gene linking these conditions with lung function. Our study revealed a shared genetic architecture between female reproductive disorders and lung function traits, providing a basis for further mechanistic investigations and potential clinical evaluation. Furthermore, our findings suggest that shared fibroproliferative and hormone-responsive pathways may offer insights into the biological mechanisms underlying these conditions.

Female

Unraveling a Diagnostic Enigma: A TECPR2 Case Solved Through Multi-Omic Genomics.

TECPR2 is a key regulator of autophagy, encoded by the TECPR2 gene. Pathogenic variants in this gene have been linked to a rare hereditary sensory and autonomic neuropathy with intellectual disability (HSAN9). We report a teenage female with a syndromic intellectual disability disorder associated with neuromuscular abnormalities. Multi-omics analysis including genomics, transcriptomics, and proteomics, together with muscle biopsy from the affected individual, were used in this clinical case. Through trio exome sequencing we identified two heterozygous variants in the TECPR2 gene, NM_014844.4: c.480G>A; p.(Gln160=) and c.2846C>A; p.(Ala949Glu). Both were classified as variants of uncertain significance due to the lack of supporting evidence for pathogenicity. Subsequent long-read sequencing phased the variants and confirmed they were in trans. Additional functional studies using RNAseq and proteomics analyses verified the pathogenicity of the variants. This case study demonstrated the value of a multi-omics assisted analysis, which complemented the traditional phenotype-first approach in reaching a definitive clinical diagnosis.

Humans

Predictive evolutionary genomics: principles, validation, and practice.

Climate change and habitat loss are driving rapid evolutionary responses in populations world-wide, which creates an urgent need for evolutionary forecasting in conservation and agriculture. Such forecasting can be categorized into three time scales: trait-based models that use multivariate quantitative genetic equations to project correlated phenotypic responses up to c. 20 generations, allele-based analyses that model allele frequency dynamics up to 100 generations, and composite adaptation scores that aggregate many small effects to yield predictions across longer horizons. However, these approaches have remained largely disconnected. Here, we present a Bayesian framework that integrates these three complementary approaches for evolutionary prediction. Our framework combines genomic, phenotypic, and environmental data to yield probabilistic predictions with explicit uncertainty. We show how predictive evolutionary forecasts can be validated with experimental evolution, field experimentation, historical specimens, and reciprocal transplants. These validated forecasts can help advance conservation and agricultural programmes by helping predict which populations are at risk of future extinction, optimizing breeding programmes for future climates, and planning ecosystem management under environmental change. By supporting a shift towards more predictive approaches in evolutionary biology, this framework may help improve our ability to manage biodiversity and food security in a changing world.

Genomics

Comparative genomic and proteomic analysis reveals orthogroup structured evolution of tick protease inhibitors.

Protease inhibitors (PIs) play central roles in regulating endogenous proteolysis and host-parasite interactions in ticks. However, the evolutionary architecture underlying their diversification across tick lineages remains insufficiently resolved. Here, we performed a genome-wide comparative analysis of predicted proteomes from 14 tick species to systematically characterize PI repertoires. In total, 4931 putative PIs were identified and grouped into 20 families using the MEROPS classification system. Further, PI families such as Antistasin, WAP-type, and Pacifastin, which have not previously been systematically reported in tick genomes, were classified. Orthogroup inference demonstrated that PI expansion is structured at the level of evolutionary lineages rather than uniformly across families. By stratifying orthogroups according to duplication burden and taxonomic conservation, we identified a broadly conserved single-copy core under strong purifying selection. Motif level analysis of serpin reactive center loops further revealed conservation of inhibitory specificity within single copy orthogroups and diversification of key functional residues in duplication-associated lineages. Integration of secretion prediction and tissue-resolved proteomics from Hyalomma anatolicum and Rhipicephalus microplus demonstrated that evolutionary stratification is reflected at the protein level. Together, these findings provide an orthogroup-resolved evolutionary framework linking duplication dynamics, molecular evolution, and tissue-level protein deployment. This integrative approach offers a systematic basis for prioritizing conserved and diversified PI lineages for future functional and anti-tick intervention studies.

Animals

CNOT1 is a potential YTHDF2 target that orchestrates maternal mRNA decay and zygotic genome activation during goat embryogenesis.

Timely and efficient degradation of maternal mRNA is essential for early embryonic development, which occurs from fertilization through the initiation of zygotic genome activation (ZGA). Yet, the regulatory mechanisms governing this process remain poorly characterized. In the present study, we investigated the function of CCR4-NOT transcription complex subunit 1 (CNOT1) during goat embryogenesis. We found that CNOT1 was upregulated during mammalian ZGA, and that its knockdown led to developmental arrest and a marked reduction in blastocyst formation. Moreover, CNOT1 knockdown impaired nascent RNA activity, resulting in 814 upregulated and 1014 downregulated genes, which were enriched for RNA splicing, regulation of chromosome organization, and RNA localization. RNA splicing analysis revealed differential splicing events in 2959 genes, of which 259 were downregulated following CNOT1 knockdown. Notably, CNOT1 was predicted to crosstalk with the m6A reader YTHDF2. Knockdown of YTHDF2 resulted in CNOT1 downregulation at the 8-cell stage in goats and increased transcription levels around polyadenylation sites during ZGA in mice. Together, these findings indicate that CNOT1 is a potential YTHDF2 target that orchestrates maternal mRNA decay and ZGA during goat embryogenesis. Our work provides new insight into the complex regulatory landscape underlying ZGA and may inform strategies to improve the efficiency of goat embryogenesis.

Animals

Emerging Principles in Spatial Functional Genomics.

Spatial transcriptomic and proteomic atlases have enabled mapping of gene programs within intact tissues, but these measurements remain largely descriptive and do not define the mechanisms controlling tissue biology. Pooled CRISPR screening provides scalable causal interrogation of gene function but remains largely confined to dissociated systems that lack spatial context. In vivo spatial functional genomics (SFG) bridges these approaches by integrating genetic perturbations with in situ transcriptomic and proteomic readouts to measure gene function within intact tissue ecosystems. By preserving spatial organization, SFG enables interpretation of perturbations through effects on cell-cell interactions, diffusible signals, multicellular niches, and tissue architecture. Here, we outline key design axes of SFG: perturbation strategy, barcoding strategy, and phenotypic readout. We discuss computational challenges, including spatial autocorrelation, neighborhood dependence, and context-aware null modeling, and highlight how SFG reveals non-cell-autonomous, architecture-dependent mechanisms of gene function, advancing toward predictive models of tissue organization and gene function.

Genomics

Systematic modular engineering of genome-integrated Escherichia coli MG1655 for high-level 2'-fucosyllactose production.

2'-Fucosyllactose (2'-FL), the most abundant human milk oligosaccharide (HMO), has attracted considerable interest for its prebiotic and immunomodulatory functions, with broad applications in infant nutrition. In this study, we report the development of a high-yield, genome-integrated 2'-FL-producing strain based on Escherichia coli MG1655 through systematic modular optimization. Starting from a single-copy BKHT strain (MGC06), we first optimized the copy number of the α-1,2-fucosyltransferase (α-1,2-FT) gene BKHT. Subsequently, the GDP-L-fucose supply was enhanced through coordinated genomic integration of the gene clusters cpsG-cpsB and gmd-fcl, while the multidrug efflux transporter gene mdfA was integrated to improve product export and strain robustness. BKHT copy number was then re-evaluated in the optimized background, with four copies yielding the highest production. The final engineered strain, harboring all genetic modifications stably integrated into the chromosome, produced 17.18 g/L 2'-FL in shake-flask culture. In fed-batch fermentation using a 5-L bioreactor, this strain achieved a titer of 154.12 g/L after 60 h, with a productivity of 2.57 g/L/h. Notably, throughout the entire fermentation process, no antibiotics or inducers were supplemented, underscoring the genetic stability and regulatory compliance of this plasmid-free system. To our knowledge, this represents the highest 2'-FL titer reported to date, positioning our engineered strain as a promising candidate for commercial 2'-FL production.

Escherichia coli

Genomic insights into end-use grain quality and nutritional traits of an ancient Indian dwarf wheat ( Triticum sphaerococcum Percival) population using a multi-locus genome-wide association study.

BACKGROUND: Triticum sphaerococcum, an ancient hexaploid wheat species, is renowned for its stress resilience and superior nutritional quality. A panel of 116 T. sphaerococcum accessions (the largest known collection at a single site globally), with six bread wheat released varieties, was evaluated for its potential for genetic quality improvement. Field experiments were conducted under standard, heat and moisture-deficit conditions across two cropping seasons for ten grain end-use quality and nutritional traits. RESULTS: Genotypes showed highly significant differences (P ≤ 0.001) for measured traits, with high broad-sense heritability resulting from substantial genotypic variance contributions. Triticum sphaerococcum consistently outperformed T. aestivum across environments, with moisture-deficit stress proving more detrimental to quality parameters than heat stress, while micronutrient content increased under stressed conditions. Trait correlations revealed that the gluten index (GI) correlated negatively with the grain hardness index (GHI), wet gluten (WG), and water-binding capacity (WB), while positively correlating with dry gluten (DG) and protein content (PRO), whereas grain iron (GFE), zinc (GZN), and protein showed consistent positive interrelationships. Two superior accessions, PAUTS10 (WG 35.13%, DG 13.71%, PRO 16.42%, GZN 50.89 ppm) and Sonamoti (WG 33.33%, DG 12.92%, PRO 16.27%, GZN 56.03 ppm), were identified, surpassing the best check variety HD3226 for quality and nutritional parameters. Multi-locus genome-wide association studies identified 30 stable quantitative trait nucleotides across environments, with candidate gene analysis revealing genes involved in transcription regulation, biosynthetic processes, metal ion homeostasis, and transport. CONCLUSIONS: Triticum sphaerococcum demonstrated superior grain quality and micronutrient potential compared with modern wheat, highlighting its value as a genetic resource for biofortification. The identification of elite accessions and stable quantitative trait nucleotides (QTNs) provides useful targets for breeding programs aimed at improving protein and micronutrient content. Integrating ancient germplasm with modern genomic tools can accelerate the development of nutritionally enhanced wheat varieties. © 2026 Society of Chemical Industry.

Triticum

Genome-wide identification of the HSP70 superfamily in tropical sea cucumber Stichopus monotuberculatus and their expression analysis under low-salinity stress.

Heat shock proteins (HSPs) are a group of evolutionarily conserved molecular chaperones that serve as indispensable core regulators in preserving cellular homeostasis and orchestrating organismal stress responses. The tropical sea cucumber Stichopus monotuberculatus, a high-value aquaculture species, is sensitive to fluctuations in environmental salinity-a challenge that has emerged as a critical bottleneck limiting its large-scale commercial cultivation. However, no systematic investigation has been conducted to characterize the HSP70 superfamily in S. monotuberculatus and elucidate its functional roles in salinity adaptation. In the present study, we performed a comprehensive genome-wide scan and identified 19 HSP70 superfamily genes in the S. monotuberculatus genome, with the HSP70IV subfamily showing remarkable gene expansion, containing 8 distinct copies. Phylogenetic analysis, conserved motif identification, and gene structure characterization demonstrated high evolutionary conservation within each HSP subfamily. These genes were unevenly distributed across the chromosomes of S. monotuberculatus, and prediction of cis-acting elements revealed that their upstream regulatory regions were enriched with numerous functional elements associated with stress response and immune regulation. Salinity stress experiments revealed that under severe low-salinity conditions (18‰), the expression levels of SmHSPA14L and multiple HSP70IV subfamily members were significantly elevated, while SmHYOU1D was significantly downregulated; in contrast, only subtle changes were detected in the expression of most HSP70 genes under moderate low-salinity stress (24‰). These findings strongly suggest that HSP70 genes, particularly the expanded HSP70IV subfamily, may act as key modulators in the low-salinity stress response. This work provides valuable insight into the molecular mechanisms underlying salinity adaptation in tropical sea cucumbers.

Animals