Search PubMedSearch

SEARCH · Search PubMed

Results for “genome evolution”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

549 records · Page 9Linked to original sources

Systematic modular engineering of genome-integrated Escherichia coli MG1655 for high-level 2'-fucosyllactose production.

2'-Fucosyllactose (2'-FL), the most abundant human milk oligosaccharide (HMO), has attracted considerable interest for its prebiotic and immunomodulatory functions, with broad applications in infant nutrition. In this study, we report the development of a high-yield, genome-integrated 2'-FL-producing strain based on Escherichia coli MG1655 through systematic modular optimization. Starting from a single-copy BKHT strain (MGC06), we first optimized the copy number of the α-1,2-fucosyltransferase (α-1,2-FT) gene BKHT. Subsequently, the GDP-L-fucose supply was enhanced through coordinated genomic integration of the gene clusters cpsG-cpsB and gmd-fcl, while the multidrug efflux transporter gene mdfA was integrated to improve product export and strain robustness. BKHT copy number was then re-evaluated in the optimized background, with four copies yielding the highest production. The final engineered strain, harboring all genetic modifications stably integrated into the chromosome, produced 17.18 g/L 2'-FL in shake-flask culture. In fed-batch fermentation using a 5-L bioreactor, this strain achieved a titer of 154.12 g/L after 60 h, with a productivity of 2.57 g/L/h. Notably, throughout the entire fermentation process, no antibiotics or inducers were supplemented, underscoring the genetic stability and regulatory compliance of this plasmid-free system. To our knowledge, this represents the highest 2'-FL titer reported to date, positioning our engineered strain as a promising candidate for commercial 2'-FL production.

Escherichia coli

Genomic characterization of a hypervirulent Aeromonas veronii NN0115 from Nile tilapia and head kidney transcriptome of infected fish reveals B-cell-dominated immune response with specific immunoglobulin downregulation.

Aeromonas veronii is a pathogen of multiple fish species, yet systematic understanding of its infection in Nile tilapia (Oreochromis niloticus) remains limited. A dominant strain, NN0115, was isolated from a natural outbreak and identified as A. veronii by 16S rRNA and whole-genome average nucleotide identity (ANI, 96.33%). Experimental infection revealed high virulence (LD50 = 3.41 × 106 CFU/mL, equivalent to 8.53 × 104 CFU/fish). The genome is 4.58 Mb (58.57% GC) and encodes 4216 proteins. Virulence factor analysis identified 1253 genes, dominated by motility-related (264) and immune modulation (208) factors. Genomic island GI2 harbors 7 virulence genes and two dual-function resistance-virulence genes. The strain is resistant to 9 of 25 agents tested but carries three RND efflux pump genes whose predicted resistance was not phenotypically observed. The head kidney transcriptome of tilapia at 24 h post-bacterial infection identified 773 differentially expressed genes; among them, 57 were immunoglobulin (Ig) genes, and 56 were down-regulated. Integration of published single-cell transcriptomic data showed that non-Ig B-cell marker genes were down-regulated by 32%, whereas Ig genes were reduced by 63%, indicating selective transcriptional suppression of Ig genes rather than a general decrease in B-cell transcriptional activity. Together, this study provides a comprehensive characterization of a highly virulent A. veronii from Nile tilapia and reveals that selective downregulation of B-cell Ig genes is the dominant transcriptional feature of the host head kidney response.

Animals

Genomic insights into end-use grain quality and nutritional traits of an ancient Indian dwarf wheat ( Triticum sphaerococcum Percival) population using a multi-locus genome-wide association study.

BACKGROUND: Triticum sphaerococcum, an ancient hexaploid wheat species, is renowned for its stress resilience and superior nutritional quality. A panel of 116 T. sphaerococcum accessions (the largest known collection at a single site globally), with six bread wheat released varieties, was evaluated for its potential for genetic quality improvement. Field experiments were conducted under standard, heat and moisture-deficit conditions across two cropping seasons for ten grain end-use quality and nutritional traits. RESULTS: Genotypes showed highly significant differences (P ≤ 0.001) for measured traits, with high broad-sense heritability resulting from substantial genotypic variance contributions. Triticum sphaerococcum consistently outperformed T. aestivum across environments, with moisture-deficit stress proving more detrimental to quality parameters than heat stress, while micronutrient content increased under stressed conditions. Trait correlations revealed that the gluten index (GI) correlated negatively with the grain hardness index (GHI), wet gluten (WG), and water-binding capacity (WB), while positively correlating with dry gluten (DG) and protein content (PRO), whereas grain iron (GFE), zinc (GZN), and protein showed consistent positive interrelationships. Two superior accessions, PAUTS10 (WG 35.13%, DG 13.71%, PRO 16.42%, GZN 50.89 ppm) and Sonamoti (WG 33.33%, DG 12.92%, PRO 16.27%, GZN 56.03 ppm), were identified, surpassing the best check variety HD3226 for quality and nutritional parameters. Multi-locus genome-wide association studies identified 30 stable quantitative trait nucleotides across environments, with candidate gene analysis revealing genes involved in transcription regulation, biosynthetic processes, metal ion homeostasis, and transport. CONCLUSIONS: Triticum sphaerococcum demonstrated superior grain quality and micronutrient potential compared with modern wheat, highlighting its value as a genetic resource for biofortification. The identification of elite accessions and stable quantitative trait nucleotides (QTNs) provides useful targets for breeding programs aimed at improving protein and micronutrient content. Integrating ancient germplasm with modern genomic tools can accelerate the development of nutritionally enhanced wheat varieties. © 2026 Society of Chemical Industry.

Triticum

Genome-wide identification of the HSP70 superfamily in tropical sea cucumber Stichopus monotuberculatus and their expression analysis under low-salinity stress.

Heat shock proteins (HSPs) are a group of evolutionarily conserved molecular chaperones that serve as indispensable core regulators in preserving cellular homeostasis and orchestrating organismal stress responses. The tropical sea cucumber Stichopus monotuberculatus, a high-value aquaculture species, is sensitive to fluctuations in environmental salinity-a challenge that has emerged as a critical bottleneck limiting its large-scale commercial cultivation. However, no systematic investigation has been conducted to characterize the HSP70 superfamily in S. monotuberculatus and elucidate its functional roles in salinity adaptation. In the present study, we performed a comprehensive genome-wide scan and identified 19 HSP70 superfamily genes in the S. monotuberculatus genome, with the HSP70IV subfamily showing remarkable gene expansion, containing 8 distinct copies. Phylogenetic analysis, conserved motif identification, and gene structure characterization demonstrated high evolutionary conservation within each HSP subfamily. These genes were unevenly distributed across the chromosomes of S. monotuberculatus, and prediction of cis-acting elements revealed that their upstream regulatory regions were enriched with numerous functional elements associated with stress response and immune regulation. Salinity stress experiments revealed that under severe low-salinity conditions (18‰), the expression levels of SmHSPA14L and multiple HSP70IV subfamily members were significantly elevated, while SmHYOU1D was significantly downregulated; in contrast, only subtle changes were detected in the expression of most HSP70 genes under moderate low-salinity stress (24‰). These findings strongly suggest that HSP70 genes, particularly the expanded HSP70IV subfamily, may act as key modulators in the low-salinity stress response. This work provides valuable insight into the molecular mechanisms underlying salinity adaptation in tropical sea cucumbers.

Animals

Regional genomic analysis of lineage distribution and transferable multidrug resistance among chicken-associated Salmonella Kentucky isolates in China.

Salmonella enterica serovar Kentucky is an important multidrug-resistant foodborne pathogen in the poultry meat supply chain. Although recent broader genomic studies have elucidated the population structure and epidemiological significance of major lineages in China (e.g., ST198 and ST314), the regional dynamics within local poultry supply chains remain insufficiently characterized. In this study, 31 chicken meat-derived isolates from Shanghai and 39 publicly available genomes from China were analyzed using antimicrobial susceptibility testing, whole-genome sequencing, phylogenetic analysis, conjugation experiments, and complete sequencing of representative plasmids. This enabled a systematic characterization of the molecular epidemiological features of the population and the mechanisms underlying resistance dissemination. Population genomic analysis revealed a lineage composition markedly different from the global epidemiological pattern: ST314 was the predominant sequence type among the Shanghai chicken-derived isolates (74.2%), whereas the internationally recognized high-risk clone ST198 accounted for only 25.8% of the local isolates. However, risk stratification analysis indicated that although ST198 was detected less frequently, it carried a significantly greater burden of acquired resistance genes and therefore represented a higher-risk resistant lineage. Functional and structural validation further elucidated the molecular basis of resistance dissemination within this high-risk lineage. Conjugation experiments confirmed the co-transfer of a multidrug resistance module carrying blaTEM-1 and blaCTX-M-267 to the recipient strain Escherichia coli J53. Complete plasmid analysis revealed that these two β-lactam resistance genes were co-localized on a 242-kb transferable plasmid flanked by Tn1331, Tn3, and multiple transposase-associated elements, thereby providing a structural basis for their horizontal transfer. This study provides important molecular epidemiological evidence for lineage-specific surveillance and risk-stratified control of resistant Salmonella in the poultry meat supply chain and further underscores the need for continuous monitoring of mobile genetic elements within a One Health framework.

Animals

Machine learning-ready genomic biomarkers: ATF3 polymorphisms predict postoperative analgesic demand through AI-compatible phenotyping.

PURPOSE: To determine whether ATF3 polymorphisms can serve as genetic biomarkers for machine learning-based precision analgesia by establishing a genotype-phenotype association suitable for predictive modeling of postoperative opioid requirements. METHODS: In a prospective cohort of 167 adults undergoing abdominal surgery, ATF3 SNPs rs3122721 and rs3125293 were genotyped. A structured dataset architecture was developed to represent genetic profiles as input features for supervised learning models, enabling translational analysis of genotype‑dependent opioid consumption over 72 h. RESULTS: Patients with homozygous genotypes of the ATF3 SNPs had significantly higher opioid requirements than non‑carriers, despite reporting similar subjective pain scores. This consistent genotype‑dependent pattern provided a clinically relevant phenotype suitable for integration into predictive algorithms. CONCLUSION: ATF3 genotyping offers a promising biomarker for computationally informed precision analgesia. By linking genomic variability to clinically meaningful outcomes within a structured clinical and genomic framework, this approach supports the future development of risk-stratified clinical decision-support systems to optimize postoperative pain management.Trial registration ChiCTR1900021991, registered 30 April 2019. SUPPLEMENTARY INFORMATION: The online version contains supplementary material available at https://doi.org/10.1007/s13755-026-00480-9.

ATF3

Genomic and One Health insights into Vibrio parahaemolyticus from environmental, seafood and clinical sources.

Vibrio parahaemolyticus is a leading cause of seafood-borne gastroenteritis worldwide, with climate warming facilitating its spread to high-latitude areas. In this study, we analyzed 212 genomes of environmental and seafood-associated isolates collected from seven cities in Zhejiang Province, China (2019-2024), alongside 228 clinical genomes from public databases. The 212 isolates were assigned to 172 sequence types (STs), with ST490 being the most frequent (5/212, 2.36%). Forty-four serotypes were identified, dominated by OL3:KUT (12.68%). High ST and serotype diversity were observed across different sample types and sources, with median pairwise single nucleotide polymorphisms (SNPs) ranging from 57,431 to 58,378, indicating comparable genetic diversity across groups. All isolates carried tlh and T3SS1 but lacked tdh and T3SS2. Resistance rates against ampicillin and cefazolin were 54.72% (116/212) and 44.34% (94/212), respectively, with multidrug resistance (MDR) detected in nine isolates, predominantly from seafood (7/9). A total of 63 distinct antimicrobial resistance genes (ARGs) spanning seven classes were identified. Isolates from aquaculture farms and wet markets exhibited greater resistance category diversity and higher ARG carriage than those from coastal or riverine sites. In contrast, the 228 clinical isolates harbored only 25 ARGs across two classes, with a significantly lower proportion of isolates carrying multiple ARG classes (0.44% vs. 6.13%, P&#xa0;<&#xa0;0.001). Human isolates formed tighter phylogenetic clusters, although a minority were closely related to environmental/foodborne strains. Overall, our findings demonstrate the genetic diversity and resistance potential of V. parahaemolyticus across environmental, seafood, and clinical sources, highlighting the importance of the One Health approach to comprehensive public health risk assessment.

Vibrio parahaemolyticus

Genomic epidemiology of clinically critical antibiotic resistance in Salmonella enterica causing bloodstream infections across six Chinese provinces, 1994-2023.

Clinically critical antibiotic-resistant Salmonella enterica (S. enterica) causing bloodstream infections remains a public health challenge. Here, we aim to reveal the emergence and trends of clinically important antibiotic resistance in S. enterica causing bloodstream infections using 833 isolates from six Chinese provincial-level administrative areas during 1994-2023. We identified 48 serovars and 64 sequence types (STs). Overall, 8.52% of 833 isolates were resistant or had decreased susceptibility to ciprofloxacin, 4.32% and 6.84% reported resistance or decreased susceptibility to third- and fourth-generation cephalosporins (3GCs and 4GCs), 1.80% reported resistance to fosfomycin, and 2.16% reported resistance to azithromycin. Across these six regions, azithromycin and fosfomycin resistance is increasing, as is decreased susceptibility or resistance to ciprofloxacin, 3GCs, and 4GCs, especially among younger children and elderly people. Clinically prioritized antibiotic resistance also varies by region, serovar, and age group. S. Paratyphi A genotype 2.3.3 strains are mainly divided into 2 lineages distributed in Guangxi and Shanghai. Within the scope of this passive surveillance dataset, S. Typhi genotype 4.3.1.2.1 was identified as the earliest documented case among the collected isolates. Our retrospective and longitudinal genomic epidemiology study provides critical data for the formulation of treatment guidelines and policies for bloodstream infections and for the monitoring and control of antimicrobial resistance.

Humans

Genome-wide characterization of the TGF-&#x3b2; superfamily identifies bmp15, gdf9, and gsdf as sex-biased candidate regulators of gonadal differentiation in the synchronous hermaphrodite Plectropomus leopardus.

The transforming growth factor-&#x3b2; (TGF-&#x3b2;) superfamily plays conserved roles in vertebrate reproduction and gonadal sex differentiation. However, its genomic repertoire and sex-biased expression patterns remain unclear in the leopard coral grouper (Plectropomus leopardus), a species with synchronous hermaphroditism. Here, we performed a genome-wide identification of the TGF-&#x3b2; superfamily, identifying 42 genes from the chromosome-level genome. Phylogenetic and synteny analyses indicated that segmental duplication under purifying selection contributed to family expansion. Expression profiling across multiple tissues and four gonadal developmental stages (undifferentiated, 120 dph; early differentiated, 15&#xa0;months; mature testis, 3&#xa0;years; mature ovary, 3&#xa0;years) identified eight gonad-enriched genes, among which bmp15 and gdf9 exhibited pronounced female-biased expression, with transcripts localized exclusively to the oocyte cytoplasm, particularly in stage II-III oocytes. In contrast, gsdf showed male-biased expression and was localized in spermatogenic cells of the testis. These reciprocal expression patterns indicate that bmp15/gdf9 and gsdf are candidate factors associated with gonadal sex differentiation. Our study provides the first comprehensive characterization of the TGF-&#x3b2; superfamily in P. leopardus and highlights bmp15, gdf9, and gsdf as candidate sex-differentiation factors in this hermaphroditic species.

Animals

In vitro evaluation of sacituzumab govitecan in non-small cell lung cancer with actionable genomic alterations.

PURPOSE: The TROP2-directed antibody-drug conjugate sacituzumab govitecan (SG) has shown substantial therapeutic benefit in several malignancies; however, preclinical evidence supporting its activity in non-small cell lung cancer (NSCLC) is rare. MATERIALS AND METHODS: We evaluated 16 NSCLC cell lines harboring actionable genomic alterations for TROP2 expression and treated them with SG or its unconjugated payload, SN-38, for 3 days to determine cytotoxic effects. Apoptosis and DNA damage signaling were assessed using flow cytometry and western blot. SG internalization and lysosomal trafficking were visualized by confocal microscopy. RESULTS: SG had greater cytotoxic potency than SN-38, across all NSCLC cell lines, independent of genomic subtype or TROP2 expression level. Cell lines that were sensitive to SN-38 showed enhanced vulnerability to SG (P < 0.0001). Higher SLFN11 expression, a recognized determinant of SN-38 responsiveness, correlated with lower SG IC50 values. Both SG and SN-38 triggered apoptotic and DNA damage responses within 6-48 h, with SG inducing stronger activation of these pathways than SN-38. SG was efficiently taken up in CUTO17 and SNU-3173 adenocarcinoma cells, with more than 60% of the conjugate internalized within 3 h and subsequently localized to lysosomes. CONCLUSION: Our study provides in vitro evidence supporting the potential activity of SG in NSCLC with actionable genomic alterations. The efficacy of SG closely paralleled intrinsic sensitivity to the SN-38 payload, suggesting that DNA-damage responses, rather than oncogenic drivers, predominantly contribute to SG activity.

Actionable genomic alterations

The cellular protein TIAR mediates rapid initiation of West Nile virus genome RNA synthesis.

During the intracellular replication cycle of West Nile virus (WNV), genome RNA synthesis is initially inefficient but increases exponentially as viral replication complexes are sequestered in invaginations in the endoplasmic reticulum. In this study, we investigated the functional role of the cellular protein TIAR (T-cell intracellular antigen-related protein) in the transcription of WNV genome RNA. Close colocalization of cytoplasmic TIAR with viral double-stranded RNA was detected by a proximity ligation assay in WNV-infected cells. TIAR binds specifically to the WNV 3'(-) SL but not to the complementary WNV 5'(+) SL in in vitro RNA binding assays. Only the 3' end of the WNV minus-strand RNA was enriched by immunoprecipitation of infected cell lysates with anti-TIAR antibody. Stable overexpression of TIAR in clonal A549 cells increased the ratio of intracellular viral plus-strand to minus-strand RNA in a dose-dependent manner. TIAR contains three RNA recognition motifs (RRMs). Biophysical data indicated that only RRM2 directly contacts RNA and that up to three TIAR molecules can bind cooperatively to the WNV 3'(-) SL RNA. These data provide additional evidence that TIAR functions as a proviral host factor facilitating exponential amplification of WNV genome production in infected cells.IMPORTANCEWest Nile virus (WNV) is a mosquito-borne orthoflavivirus associated with increasing global human disease incidence. The molecular mechanisms underlying viral replication are not fully understood. In early stages of infection, viral genome transcription is inefficient; however, in late stages, viral genome transcription increases exponentially. T-cell intracellular antigen-related (TIAR) protein is a cellular protein that has been shown to interact with the 3' end of the WNV negative-sense antigenomic RNA. We obtained data showing colocalization of cellular TIAR with viral replication complexes in infected cells and an increased ratio of intracellular genomic to antigenomic viral RNA in TIAR-overexpressing cells, and confirmed preferential binding of TIAR to the 3' end of the WNV antigenome both in vitro and in infected cell extracts. We also demonstrated that multiple TIAR proteins can bind cooperatively to the WNV 3'(-) stem-loop RNA. These data provide supporting evidence for a model of TIAR-mediated rapid initiation of nascent genome RNA synthesis in infected cells.

TIAR

PGR expression as a pharmacogenomic companion biomarker to GENE70-derived genomic risk in ER-positive/HER2-negative breast cancer.

BACKGROUND: The biology of the estrogen receptor-positive (ER+) and human epidermal growth factor receptor 2-negative (HER2-) breast cancers is heterogeneous even when they are categorized by their risk via genomics. Transcriptomic PGR expression reflects endocrine pathway activity and may provide complementary biological information within established GENE70-derived genomic-risk categories. Whether this molecular marker improves the biological interpretation of genomic-risk stratification beyond conventional clinicopathological assessment remains uncertain. OBJECTIVES: The aim of this study was to determine whether transcriptomic PGR expression provides complementary biological and prognostic information within reconstructed GENE70-derived genomic-risk categories and refines the characterization of endocrine-related tumour biology in ER-positive/HER2-negative breast cancer. METHODS: This study analysed publicly available transcriptomic and clinical data from three cohorts: METABRIC (discovery cohort), GSE96058/SCAN-B cohort (validation cohort) and TCGA-BRCA cohort (molecular validation cohort). The GENE70-derived genomic-risk score was reconstructed for each cohort using matched genes. Cox regression, Kaplan-Meier analysis and subgroup comparisons were used to assess relationships between PGR expression, clinicopathologic variables, molecular features and survival outcomes. RESULTS: Across the three independent cohorts, low transcriptomic PGR expression was consistently associated with higher GENE70-derived genomic risk, increased MKI67 expression, reduced ESR1 expression and enrichment of the Luminal B subtype. Survival findings differed between cohorts. In the discovery METABRIC cohort, transcriptomic PGR expression showed heterogeneous associations with survival, particularly within GENE70-derived high-risk subgroups, whereas the external GSE96058/SCAN-B validation cohort demonstrated consistent associations between low PGR expression and poorer overall survival in both the overall ER-positive/HER2-negative population and GENE70-derived high-risk subgroups. CONCLUSION: These findings suggest that transcriptomic PGR provides complementary biological and prognostic information within GENE70-derived genomic-risk categories. However, because treatment response was not evaluated in the present study, the findings should not be interpreted as evidence of predictive or pharmacogenomic utility and prospective studies incorporating treatment-response analyses are required before such applications can be established.

Humans

Genome-wide identification and expression analysis of the CREB/ATF family and its potential role in melanogenesis in the Manila clam (Ruditapes philippinarum).

Ruditapes philippinarum is an economically important bivalve species in China, and shell color is a trait of ecological and commercial significance. Melanin is a key determinant of shell color, and members of the CREB/ATF family have been reported to participate in melanogenesis in other organisms. In this study, members of the CREB/ATF family were systematically identified at the whole-genome level based on genomic and transcriptomic datasets, followed by analyses of their phylogenetic relationships, gene structures, and expression patterns. A total of six CREB/ATF family members were identified and classified into five subfamilies. Expression profiling and RT-qPCR validation revealed that most CREB/ATF genes were highly expressed in the mantle and displayed clear differences among shell-color phenotypes. Except for RpATF4 and RpCREBZF, most members exhibited relatively high expression levels in dark-colored shell strains, particularly in black and zebra-striped clams. Moreover, most genes showed low expression during early embryonic and larval stages but increased expression at the single-siphon spat and juvenile stages. These results suggest that the CREB/ATF family may be involved in melanin-associated shell-color regulation in R. philippinarum, providing important candidate genes and a theoretical basis for further elucidating the molecular mechanisms of shell-color formation in mollusks.

Animals

Multimodal alignment improves generalizability of genomic biomarker prediction in computational pathology.

Computational pathology models that use digitized histopathology whole-slide images have the potential to become a cost-effective and scalable alternative to molecular assays for the prediction of genomic biomarkers, a key task in precision oncology. However, as new genomic biomarkers are discovered or quantified, large, labeled datasets must be prospectively collected to train new models. To address this challenge, we developed multimodal alignment for biomarker learning and generalization (MARBLE), a multimodal contrastive pretraining strategy that integrates structured biomarker knowledge into representation learning of histopathology images. MARBLE aligns histopathology-derived representations with representations of genomic biomarkers generated by a large language model (LLM) and a protein language model (PLM). This biologically informed alignment enables data-efficient generalization to novel, out-of-distribution biomarkers. Using the MSK-IMPACT cohort of over 40,000 patients across multiple biomarker panel versions, we design experiments grounded in real-world data to demonstrate the value of our proposed approach.

CP: computational biology

Integrative genomic and transcriptomic analyses identify key regulators of skin pigmentation in Larimichthys crocea.

The yellow body coloration of large yellow croaker (Larimichthys crocea) constitutes a crucial economic trait, yet its underlying genetic regulatory mechanisms remain poorly understood. This study systematically elucidated the molecular basis of body color variation by integrating genome resequencing and skin transcriptome analyses, combined with the contextual analysis of key pigmentation-related genes and phenotypic histological validation. 200 phenotyped individuals (including yellow-selected lines, F1 progeny, and normal control groups, all derived from a well-characterized aquaculture stock) identified 39 significantly associated SNPs (-log&#x2081;&#x2080;(P)&#xa0;&#x2265;&#xa0;6), mapping to multiple candidate genes. These genes were significantly enriched in pathways related to pigment deposition (GO:0033059), melanosome organization (GO:0032438), melanogenesis, and tyrosine metabolism. Cross-developmental stage transcriptome analysis revealed 2395 differentially expressed genes (DEGs). Multi-omics integration identified eight overlapping candidate genes, including tyrp1, slc45a2, oca2, and dgat2, among which tyrp1 was prioritized for in-depth validation based on its core regulatory role in eumelanin synthesis, significant SNP association signal, and consistent downregulation in transcriptomic data. Experimental validation demonstrated that the g.895C&#xa0;>&#xa0;T mutation in exon 2 of tyrp1b was strongly significantly associated with the yellow phenotype: the frequency of mutant genotypes (TT/CT) reached 92.86%in the yellow-selected group, whereas the control group exclusively exhibited the wild-type genotype (CC). qPCR confirmed significantly downregulated tyrp1b expression in the skin of yellow individuals, consistent with the transcriptome trend. Histological and stereomicroscopic observations of skin tissues further validated the physiological basis of the yellow phenotype, revealing a significant reduction in melanophore number and abnormal melanosome morphology in yellow-phenotype individuals, accompanied by increased xanthophore density. These results suggest that tyrp1b mutation is strongly associated with the yellow phenotype. However, the presence of a wild-type CC individual in the yellow group indicates that this mutation is not strictly required for yellow coloration, suggesting that other genetic or environmental factors may also contribute to the phenotype, Additionally, downregulation of the carotenoid metabolism gene bco2 coupled with upregulation of xdh, together with the functional changes of slc45a2 and oca2, may synergistically promote xanthophore pigment deposition, contributing to the yellow phenotype. As melanin synthesis in large yellow croaker relies on the conserved tyrosinase pathway and transporter proteins, mutations in associated genes (tyrp1b, slc45a2, oca2) represent a primary underlying cause for the loss of melanin-based coloration and transition to a yellow phenotype in L. crocea. These findings provide key molecular targets and a theoretical foundation for molecular breeding of body color in this species, and also enrich the understanding of xanthism regulatory mechanisms in teleosts.

Animals

Genome mining of alkaliphilic cyanobacterial consortia: identification of biosynthetic gene clusters in Sodalinema and associated heterotrophs.

Alkaline soda lakes are high-pH environments that host specialized microbial communities with potential for biotechnology and natural product discovery. We characterized three Sodalinema-dominated cyanobacterial consortia enriched from Canadian soda lakes over 510 days. Using hybrid metagenomic sequencing and metatranscriptomics across pH, alkalinity, and temperature gradients, we reconstructed high-quality metagenome-assembled genomes and assessed functional activity. All consortia converged toward cyanobacteria dominance and exhibited temperature optima between 21&#xb0;C and 30&#xb0;C. Phylogenetic analysis placed Sodalinema genomes within a distinct clade affiliated with Candidatus Sodalinema alkaliphilum. Genomic analysis indicated complete biosynthetic pathways for vitamin B5, vitamin B7, and the molybdenum cofactor, but incomplete pathways for vitamins B1, B9, and B12, consistent with patterns observed in Sodalinema yuhuli. Metatranscriptomic profiles showed increased expression of genes involved in phycocyanin and carotenoid biosynthesis at pH 10.2 relative to pH 8.5. Biosynthetic gene cluster analysis revealed that most secondary metabolic potential resided in heterotrophic community members. Roseinatronobacter encoded pathways for N-acyl homoserine lactones, osmoprotectants, betalactones, and prodigiosin, while Alkalimonas, Wenzhouxiangella, and members of the Kiloniellales encoded clusters for lanthipeptides, cyclodipeptides, hydrogen cyanide, and pyrroloquinoline quinone. These findings indicate functional partitioning within the consortia and highlight the contribution of heterotrophs to secondary metabolism.IMPORTANCEAlkaline soda lakes contain microbial communities adapted to high pH that remain underexplored for biotechnology. This study focuses on Sodalinema, a filamentous cyanobacterium that dominates enriched consortia from Canadian soda lakes, and its associated heterotrophic partners. We show that while Sodalinema drives primary productivity, heterotrophic bacteria encode most of the pathways for antimicrobial and signaling compounds. These interactions may support community stability and defense against competing microorganisms. By linking genomic potential with gene expression, this work identifies alkaline cyanobacterial consortia as a source of bioactive compounds and provides a framework for exploring extremophilic microbial communities for natural product discovery.

Sodalinema

A genome-wide coverage-based pipeline for the identification of host-derived candidate DNA biomarkers from cell-free blood.

We have created a new data-analysis pipeline for the discovery of host-specific candidate DNA biomarkers derived from sequencing data of cell-free blood. Unlike approaches that rely on specific molecular or genetic signatures, our method leverages the coverage distribution of cell-free DNA sequences mapped to a reference genome, applying statistical analyses to identify informative short genomic regions for biomarker discovery. The pipeline is applicable to diverse diseases and can be used to analyze cell-free DNA sequences from plasma or serum to identify candidate biomarkers that are characteristic of disease states in mammals. Core functionalities were developed in Java and integrated with open-source software tools for the preprocessing of raw sequencing data, complemented by Python scripts for the machine-learning analysis and statistical validation. The pipeline is designed for HPC use and users can access the pipeline through a Galaxy workflow, which offers a user-friendly web interface for input selection prior to execution and analysis progress monitoring. Performance tests, carried out using duplicate sets of COVID-19 samples and controls, showed linear scalability of execution time with an increasing dataset size, as well as a substantial reduction in execution time through parallelized computation, whereby each HPC node is used to process the data of one chromosome. Further statistical tests confirmed the quality of the pipeline's results by showing that the set of identified candidate biomarkers remained stable across varying dataset sizes.

Biomarkers