Search PubMedSearch

SEARCH · Search PubMed

Results for “whole‐genome sequencing”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

550 recordsLinked to original sources

The cold case of state transition 7 (stt7) mutants of Chlamydomonas reinhardtii, solved by whole-genome sequencing.

The process of State Transitions (ST) corresponds to an STT7 kinase-driven redistribution of the transmembrane LHCII antenna proteins between Photosystem II (PSII) and Photosystem I (PSI), which results from changes in their phosphorylation state. For the past two decades, two LHCII-kinase mutants, stt7-1 and stt7-9, have been instrumental in the study of STs in Chlamydomonas reinhardtii, the former being a null mutant for the kinase but quasi-sterile in crosses, while the latter, although fertile, has a leaky phenotype. Using long-read sequencing, this study further characterized the genetic lesions of the stt7 mutant strains through whole-genome reconstruction and de novo chromosome assembly. In addition, two new stt7 null mutants were generated, one derived by crosses from the original stt7-1 and one obtained by Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR)-associated protein 9 (Cas9) technology. This work provides a comprehensive genomic characterization of the original stt7-1 null mutant, revealing extensive chromosomal rearrangements and high levels of aneuploidy, associated with increased cell size and meiotic dysfunction. Reassessment of their physiology and genetic backgrounds highlights the need for caution in interpreting genetic information. We thus produced more reliable null mutants for the LHCII-kinase, amenable to genetic crosses for the study of STs in a variety of genetic backgrounds.

Chlamydomonas reinhardtii

Molecular Diagnostics for WHO Priority Bacterial Pathogens: A Bibliometric Mapping of Diagnostic Platforms, Resistance Markers, and Antimicrobial Resistance Research Trends.

Antimicrobial resistance (AMR) constrains effective treatment and carries implications for infection control, surveillance, and public health. The World Health Organization (WHO) priority bacterial pathogen framework has intensified the need for diagnostic innovation by redefining research priorities around organisms combining high disease burden with complex resistance profiles. Molecular diagnostics have accordingly moved beyond culture-based workflows, integrating rapid pathogen identification, resistance-marker detection, genomic surveillance, and clinical decision support. The present study conducted a bibliometric mapping of the literature on WHO priority pathogens. Rather than addressing resistance at a general level or a single pathogen or technology, it integrates priority pathogens, molecular platforms, and resistance markers within a single framework, tracing their joint thematic and temporal evolution along an explicit pathogen-platform-marker axis. Scopus-indexed articles and reviews (2000-2025) were retrieved, yielding 1746 publications after screening adapted from the Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) guidelines. Analyses used Bibliometrix/Biblioshiny, R, and VOSviewer. The literature expanded markedly after 2018, led by China and the United States. Methicillin-resistant Staphylococcus aureus (MRSA), Mycobacterium tuberculosis, Enterococcus faecium, and the Enterobacterales-carbapenemase axis constituted the principal thematic cores, whereas conventional polymerase chain reaction (PCR)/nucleic acid amplification testing (NAAT) and whole-genome sequencing were the dominant platforms. Overall, the field has evolved from pathogen detection into an AMR-centered translational domain encompassing resistance prediction, genomic epidemiology, surveillance, and clinical decision support. Diagnostic development, stewardship, and surveillance depend on hybrid workflows coupling rapid marker-targeted assays with genome-based characterization, delivering actionable resistance within clinically meaningful timeframes, and extending coverage to underrepresented pathogens and platforms.

Humans

Novel Germline ELP1 Splice-Acceptor Variant in NF1-Negative Optic Pathway Glioma: Expanding the Clinical Spectrum Associated With ELP1 Variation.

We report a 7-year-old boy with NF1-negative optic pathway glioma harboring a novel germline ELP1 splice-acceptor variant (NM_003640.5:c.2205-2A>G) identified by whole-exome sequencing. The variant was likely pathogenic (ACMG/AMP: PVS1, PM2) and inherited from an asymptomatic father, consistent with incomplete penetrance, expanding the limited evidence linking germline ELP1 variation to gliomas.

Humans

Whole genome sequencing of unusual Hepatitis C virus subtypes and drug resistance analysis during direct-acting antiviral therapy in India.

INTRODUCTION AND OBJECTIVES: Pangenotypic direct-acting antivirals (DAA) are effective against highly prevalent Hepatitis C virus (HCV) subtypes, but have been clinically validated almost exclusively in high-income countries. Unusual HCV subtypes may carry natural polymorphisms, potentially impacting DAA susceptibility. We conducted full-genome characterization and resistance analysis of unusual HCV subtypes in patients receiving DAA treatment. PATIENTS AND METHODS: In this prospective hospital-based study, eligible patients were screened for anti-HCV antibodies and active infection was confirmed by diagnostic 5'NCR-based HCV RNA detection. Genotyping was performed by core region sequencing, and viral load quantified by real-time PCR. For whole genome sequencing, multiplex primers were designed using alignments of global reference sequences. Sequencing was carried out using the Oxford Nanopore Technology platform. Phylogenetic analysis used multiple sequence alignment and the HCV-GLUE resource for resistance-associated substitution (RAS) analysis. RESULTS: Predominant genotype was genotype 3 in 64.3% (n = 45); genotype 6 in 21.4% (n = 15); and genotype 1 in 14.2% (n = 10). Unusual HCV subtype 6xa was detected in two patients and showed no NS5A resistance mutations. One genotype 3b patient relapsed at 24 weeks post-DAA treatment completion and carried NS5A resistance-associated substitutions 30 K and 31 M both at baseline and at relapse, conferring high-level resistance to NS5A inhibitors. CONCLUSION: This is the first report from India of whole genome sequencing of HCV subtype 6xa. The identification of NS5A resistance mutations in the 3b relapse case underscores challenges for global HCV elimination strategies.

Humans

Evaluation of one-step amplicon-based targeted enrichment for SARS-CoV-2 whole-genome sequencing using the Midnight amplicon scheme.

Genomic surveillance proved invaluable during the COVID-19 pandemic for tracking SARS-CoV-2 variants and guiding outbreak responses, underscoring the ongoing need to reduce whole-genome sequencing (WGS) costs and improve workflow efficiency to ensure accessibility in resource limited settings. Here, we evaluated a one-step reverse transcription polymerase chain reaction (RT-PCR) approach using the Midnight V2 primer scheme for targeted amplification of the SARS-CoV-2 genome, assessed its compatibility with Illumina sequencing, and compared its performance to a well-established two-step method. Initially, we determined optimal RT-PCR reaction conditions using the Midnight V2 primer panel for the one-step RT-PCR kit and scaled reaction volumes for both RT-PCR and library preparation. Clinical specimens (n = 53) that had undergone routine WGS for surveillance purposes using the established two-step RT-PCR method were compared using the one-step RT-PCR assay. For samples with genome completeness greater than 70%, both methods gave comparable results with similar sequence coverage and 100% concordance for lineage assignment. Further investigation revealed a higher percentage of reads aligning to the SARS-CoV-2 genome with a greater depth of coverage using the one-step method compared to the two-step method. Finally, analysis of scaled one-step and library reaction volumes revealed significant cost savings for samples undergoing WGS. Overall, the results presented here verify the accuracy and reproducibility of one-step targeted amplification and offer an efficient and cost-effective workflow for routine SARS-CoV-2 genomic surveillance.

Humans

Longitudinal whole-genome analysis of bluetongue virus identifies conserved serotype-specific genomes and distinct genomic constellations within a Colorado sheep flock (2021-2023).

Bluetongue virus (BTV) is a segmented double-stranded RNA virus of ruminants transmitted by Culicoides spp. biting midges. Although the genome consists of ten segments, classification into serotypes is primarily based on genome segment 2. However, reassortment among genomic segments is a major driver of BTV evolution and diversity. This study used longitudinal whole-genome sequencing to characterize BTV genomes collected from 2021 to 2023 within a single sheep flock in Colorado, where multiple serotypes co-circulate. Whole-genome sequences were generated from fourteen blood samples representing four serotypes: BTV-6, -11, -13, and -17. Longitudinal sampling identified multiple BTV serotypes within individual sheep across consecutive years. Tanglegram analysis comparing segment phylogenies to the segment 2 tree demonstrated incongruent topologies across all genomic segments, suggestive of reassortment or the circulation of distinct genomic constellations. Nucleotide-level comparisons revealed high sequence homology among same-serotype samples from the same year, while the greatest genetic divergence was observed among BTV-17 genomes collected in different years. Additionally, all BTV-13 genomes contained a previously undescribed nonsynonymous substitution in segment 10 predicted to extend the encoded protein by three amino acids. Together, these findings demonstrate that highly conserved BTV genomes and distinct genomic constellations can be detected at the flock level across multiple years. This longitudinal whole-genome approach reveals the genetic complexity of endemic BTV populations, including novel variants and genomic patterns consistent with reassortment that are lost with conventional serotyped-based approaches, highlighting the need to integrate whole-genome characterization into endemic BTV monitoring programs.

Animals

Exploring the mechanism of aroma production in fermented cherry juice by L. brevis LD1.0600 using flavomics and whole genome analysis.

This study focused on L.brevis LD1.0600 with excellent fermentation traits: it analyzed genome-wide key regulatory genes for micro-metabolites, combined with fermented cherry juice flavor metabolomics data, and used machine learning to explore correlations between gene regulation, metabolite production, and flavor formation. The SVM model screened and verified fermented cherry juice VOCs; through OAV and flavor wheel analysis, LD1.0600 emerged as the top-performing strain, with a sweet, fruity dominant aroma. Key aroma-active components (OAV > 100) included 2-methoxy-4-vinylphenol, benzaldehyde, 2-methyl-butanoic acid and hexanoic acid, and 2-methoxy-4-vinylphenol and hexanoic acid elevated by LD1.0600-regulated genes (Chrom1-001884, Chrom1-000925, fabF and Chrom1-000199). At the same time, through research, a "strain screening-SVM screening of DVCs-OAV screening of key aroma components-whole genome sequencing of flavor regulatory genes" system was established. This system can not only be applied to the screen fermentation strains, but also can be extended to the application of other fermentation products.

Fermentation

Routine methods misidentify Serratia spp.: Limitations of MALDI-TOF MS revealed by whole-genome sequencing.

Accurate species-level identification within the genus Serratia remains challenging due to extensive phenotypic overlap and high genomic relatedness among closely related and recently described taxa. This study presents an evaluation of routine and genome-based identification approaches applied to clinical Serratia isolates, integrating phenotypic assays, MALDI-TOF MS (Bruker Daltonics), 16S rRNA gene sequencing, and Whole-Genome Sequencing (WGS). A total of 103 isolates collected from a teaching hospital were analyzed. WGS was performed on a subset of isolates. Conventional biochemical methods classified all isolates as Serratia marcescens, whereas MALDI-TOF MS identified 60.1% as S. marcescens, 11.6% as S. ureilytica, and 28.1% just at the genus level. Peak analysis from MALDI-TOF MS revealed specific peaks associated with S. marcescens and S. ureilytica, but limited discriminatory power. WGS of six isolates initially identified as S. ureilytica by MALDI-TOF MS revealed reclassification as Serratia sarumanii (n = 5) and Serratia montpellierensis (n = 1), supported by Average Nucleotide Identity (ANI), Average Amino Acid Identity (AAI), and Digital DNA-DNA Hybridization (dDDH) thresholds. In contrast, 16S rRNA analysis showed limited species-level resolution. Phylogenomic and SNP-based analyses confirmed these classifications with strong support. Overall, this study underscores the critical role of high-resolution genomic approaches for precise species identification and highlights the need for continuous expansion and curation of MALDI-TOF MS reference databases to support reliable clinical diagnostics and epidemiological surveillance of emerging Serratia species.

Spectrometry, Mass, Matrix-Assisted Laser Desorpti

Genomic characterization and pathogenicity of ruminant Listeria monocytogenes isolates in a murine oral infection model.

Listeria monocytogenes is a major foodborne pathogen; its ruminant isolates display zoonotic characteristics, causing similar clinical signs in humans, including abortion and encephalitis. However, data on whole genome sequencing and pathogenicity of ruminant L. monocytogenes isolates remain sparse. This study aimed to analyze the genotypic characteristics of L. monocytogenes isolates from ruminants with listeriosis. Furthermore, we assessed the in vivo pathogenicity of four ruminant L. monocytogenes isolates, characterized via whole-genome sequencing-based genetic clustering, in orogastrically inoculated mice. The isolate LM18 (serotype 1/2b, ST224, SL6178) had the lowest lethal dose compared to the other three isolates including previous hypervirulence type (serotype 4b, ST1, SL1) and caused secondary bacteremia in lungs, with sustained bacterial loads in the spleen and liver. Genomic (listeria pathogenicity island -1 and -3) and virulence gene (actA and llsX) mutation analyses associated with virulence suggested from well-recognized studies could not elucidate the virulence of the isolates. SSI-1, which only exists in the isolate LM18 (serotype 1/2b, ST224, SL6178), may help L. monocytogenes survive in the gastrointestinal environment, thereby affecting its virulence. Further research should investigate the role of SSI-1 in the pathogenicity of L. monocytogenes. Moreover, additional studies utilizing larger datasets of ruminant isolates are required to validate our genotypic characterization and to obtain a comprehensive picture of further genotypic differences crucial for L. monocytogenes pathogenicity.

Animals

Comparative genomic epidemiology of food- and patient-derived diarrheagenic Escherichia coli from sentinel surveillance in Southeast China.

Diarrheagenic Escherichia coli (DEC) remains an important foodborne pathogen, yet long-term comparative genomic surveillance data jointly characterizing food-derived and patient-derived isolates remain limited. This surveillance-based comparative study integrated antimicrobial susceptibility testing and whole-genome sequencing to characterize diarrheagenic Escherichia coli isolates recovered from food and patient sources in Lishui, Southeast China, during 2018-2025, with emphasis on occurrence, resistance profiles, genomic backgrounds, and plasmid replicon-associated features. Antimicrobial susceptibility testing was performed for 258 selected isolates, and whole-genome sequencing was conducted for a curated analytical subset of 204 isolates. The sequenced subset was used for diversity-oriented comparative genomic analysis rather than for unbiased prevalence estimation of the entire DEC collection. EAEC predominated in both sources, although food-associated occurrence was heterogeneous across categories, with the highest recovery rate observed in raw meat. Patient-derived isolates showed a broader overall resistance burden, whereas food-derived isolates retained substantial resistance to tetracycline, chloramphenicol, and florfenicol. Phylogenetic analysis showed partial overlap in genomic backgrounds between food-derived and patient-derived isolates, while representative resistance determinants displayed both broadly distributed and lineage-enriched patterns. Replicon-based plasmid profiling identified 42 plasmid types, including 12 detected in both sources, with IncF-related replicons predominating among these shared profiles. Several food-derived isolates carried multiple plasmid replicon types that were also observed in patient-derived isolates. Overall, food-derived and patient-derived DEC showed partial overlap in genomic backgrounds, resistance determinants, and replicon-defined plasmid profiles within this surveillance setting, while retaining source-associated heterogeneity. These findings should be interpreted as surveillance-based comparative evidence rather than as evidence of direct source attribution or transmission.

Humans

Nanopore-based epigenomic profiling reveals the absence of widespread CpG methylation in the African swine fever virus genome.

DNA methylation is a critical epigenetic mechanism implicated in regulating replication and transcription in DNA viruses. However, the epigenetic landscape of African swine fever virus (ASFV), a large double-stranded DNA virus infecting pigs, remains controversial. Here, we systematically profiled the DNA methylome of the first ASFV strain isolated in Hong Kong (HK_NT_202103) using Oxford Nanopore Technologies (ONT) R10.4.1 sequencing. We employed a paired design: native whole-genome sequencing (WGS) against a methylation-free whole-genome amplification (WGA) control. Using conservative thresholds, we found no evidence of 5-methylcytosine (5mC), especially typical CpG methylation, across the viral genome. Importantly, clear CpG methylation signals were successfully detected in the host genome from WGS data, confirming the functionality of the workflow to detect 5mC at CG sites. While widespread 5mC seems absent, a small number of putative N6-methyladenine (6mA) loci were identified. A specific 6mA candidate exhibited raw ionic current disruptions and gene-level intersection with another ASFV isolate (CAS19-01/2019), although it lacked single-base consensus across different methylation callers or between the two isolates. Although our biological findings are restricted to a single isolate under specific experimental conditions, this study introduces a novel, highly rigorous ONT framework for viral epigenomics research. Furthermore, the absence of ASFV CpG methylation indicates that host CpG-depletion remains a viable strategy for viral metagenomic enrichment. Ultimately, our work offers a critical methodological baseline for ASFV surveillance and highlights the necessity of targeted experimental validation for rare viral modifications.

African Swine Fever Virus

Whole-transcriptome RNA sequencing and ceRNA network analyses provide novel insights into the antibacterial immune response of Hippocampus abdominalis against Vibrio harveyi.

Long non-coding RNAs (lncRNAs) stand as newly-arisen molecular types that exert regulatory effects, able to operate as competitive endogenous RNAs (ceRNAs) to engage microRNAs (miRNAs) in interaction, resulting in the recovery of target mRNA expression and activity. Increasing evidences indicate that the ceRNA network affects various biological processes in mammals, including development, cellular differentiation, metabolism, immune response, and disease pathogenesis. In teleost fish, the lncRNA-miRNA-mRNA regulatory networks have been reported occasionally. However, up to now, the roles of lncRNAs in the big-belly seahorse (Hippocampus abdominalis) remains unclear. In this study, we reported for the first time, via whole-transcriptome RNA sequencing, the lncRNA mediated ceRNA regulatory network in Vibrio harveyi-infected H. abdominalis. A total of 4197 differentially expressed mRNAs (DE-mRNAs), 1317 DE-lncRNAs, and 183 DE-miRNAs were identified. Furthermore, the crosstalk between miRNAs and lncRNAs as well as between miRNAs and mRNAs was inferred based on the negative correlations between miRNAs and their target lncRNAs/mRNAs. A core immune associated lncRNA-miRNA-mRNA putative regulatory network was thus constructed, comprising 211 lncRNA-miRNA and 224 mRNA-miRNA pairs. In conclusion, our findings provide an integrative overview of the ceRNA regulatory networks on the underlying immune responses to V. harveyi infection in the big-belly seahorse, and offer a solid theoretical foundation for the comparative immunological research of teleost fish.

Animals

Diagnostic and clinical utility of exome sequencing and chromosomal microarray in children with GDD/iD: a meta-analysis.

BACKGROUND: Global developmental delay/intellectual disability (GDD/ID) is among the most common neurodevelopmental disorders, with up to half of cases are attributed to genetic factors. Chromosome microarray (CMA) has traditionally been the primary genetic test for idiopathic GDD/ID. However, whole exome sequencing (WES) and whole genome sequencing (WGS) have recently emerged, substantially increasing diagnostic yields in these populations. METHODS: We conducted a comprehensive literature search of PubMed, Scopus, EMBASE, and the Cochrane Library from inception to April 29, 2025. Studies reporting the diagnostic utility of these tests in children with GDD/ID were included and analyzed. RESULTS: A total of 102 studies, comprising 55,752 children, were reviewed. The pooled diagnostic yield of WES was 0.37 (95% CI: 0.33-0.41; I2 = 93%), significantly higher than that of CMA at 0.19 (95% CI: 0.16-0.21; I2 = 95%). Subgroup analyses showed that WES yielded significantly higher diagnostic rates than CMA in both same-sample comparisons (OR = 2.27, 95% CI: 1.08-4.78) and different-sample comparisons (OR = 1.65, 95% CI: 1.15-2.37). Only one study evaluated WGS, reporting a diagnostic yield of 0.27. Meta-regression revealed a significant association between CMA diagnostic yield and the proportion of male participants (p&#x2009;<&#x2009;0.01), but not with WES. No significant difference in diagnostic utility was observed between isolated GDD/ID and GDD/ID with comorbidities. CONCLUSION: In children with unexplained GDD/ID, WES demonstrates superior diagnostic and clinical utility compared to CMA. Incorporating WES as a first-line investigation in the diagnostic evaluation of GDD/ID may be warranted.

Humans

A genome-wide coverage-based pipeline for the identification of host-derived candidate DNA biomarkers from cell-free blood.

We have created a new data-analysis pipeline for the discovery of host-specific candidate DNA biomarkers derived from sequencing data of cell-free blood. Unlike approaches that rely on specific molecular or genetic signatures, our method leverages the coverage distribution of cell-free DNA sequences mapped to a reference genome, applying statistical analyses to identify informative short genomic regions for biomarker discovery. The pipeline is applicable to diverse diseases and can be used to analyze cell-free DNA sequences from plasma or serum to identify candidate biomarkers that are characteristic of disease states in mammals. Core functionalities were developed in Java and integrated with open-source software tools for the preprocessing of raw sequencing data, complemented by Python scripts for the machine-learning analysis and statistical validation. The pipeline is designed for HPC use and users can access the pipeline through a Galaxy workflow, which offers a user-friendly web interface for input selection prior to execution and analysis progress monitoring. Performance tests, carried out using duplicate sets of COVID-19 samples and controls, showed linear scalability of execution time with an increasing dataset size, as well as a substantial reduction in execution time through parallelized computation, whereby each HPC node is used to process the data of one chromosome. Further statistical tests confirmed the quality of the pipeline's results by showing that the set of identified candidate biomarkers remained stable across varying dataset sizes.

Biomarkers

Tigecycline-resistant Staphylococcus in waiting pens of a pig slaughterhouse: genomic insights into a food safety alert.

BACKGROUND: The waiting pens of slaughterhouses represent a critical control point in the 'farm-to-fork' continuum, yet their role in the emergence and dissemination of antimicrobial resistance remains understudied. This study investigated tigecycline-resistant Staphylococcus (TRS) in these high-risk zones to assess their prevalence, resistance mechanisms, and transmission dynamics. METHODS: 400 samples were collected from the waiting pens of a pig slaughterhouse in Guangzhou, China. Antimicrobial susceptibility testing, whole-genome sequencing, phylogenetic analysis, and molecular cloning were employed to characterize resistance mechanisms and transmission patterns. RESULTS: 78 TRS strains were isolated and classified into three species, including S. borealis, S. ureilyticus, and S. pasteuri. These isolates exhibited multidrug-resistant phenotypes and carried new mutations in rpsJ and tet(M), which were functionally confirmed to reduce tigecycline susceptibility. Phylogenetic evidence demonstrated clonal transmission between pig farms and the slaughterhouse. The tet(M) gene was located within Staphylococcal cassette chromosome mec elements mediated by IS257, while tet(L) was carried by plasmids formed through IS256/IS257-mediated recombination. CONCLUSIONS: Waiting pens serve as crucial reservoirs for the amplification and dissemination of antimicrobial resistance. Our findings underscore the urgent need for enhanced biosecurity measures, improved waste management, and routine molecular surveillance in these high-risk zones to mitigate the spread of resistance along the food production chain.

Animals

Upscaling Genotyping by Amplicon Sequencing With GBAS-GUI.

Genotyping by amplicon sequencing (GBAS) is a relatively low-cost approach for generating genotypic data compared with established genomic methods, making it highly scalable and particularly suitable for large-scale genetic monitoring projects. However, most existing analytical pipelines are either marker-specific, insufficiently scalable, or lacking efficient data management systems for the long-term integration of genotypic information, limiting the full potential of GBAS. Here, we address this gap by introducing GBAS-GUI (https://github.com/sonnenbe-dot/GBAS-GUI), a pipeline capable of generating GBAS-based genotypic data for a wide variety of loci at scale. GBAS-GUI integrates a graphical user interface with multiple checkpoints to improve accessibility and robustness. It implements multiprocessing architecture and a relational database that links genotypic data with associated sample metadata to enhance scalability and data management. The pipeline further enables marker screening through automated calculation of polymorphism information content (PIC) and implements a strategy to recover homologous genotypic information from paralogous loci with non-overlapping amplicon length ranges. Using multiple empirical datasets, we demonstrate substantial improvements in processing speed, database management and handling artefacts related to co-amplification of unspecific regions and duplicates of the same genomic region. We further show that incorporating the full sequence information captured by an amplicon increases marker information content beyond what is achievable with length-based genotyping alone and expands the analytical versatility of GBAS. Overall, GBAS-GUI provides a robust, scalable and versatile framework that unlocks the potential of GBAS for large-scale population genetic and phylogeographic studies.

Genotyping Techniques

Genomic and food-safety evaluation of Staphylococcus chromogenes in Chinese dairy milk.

Non-aureus staphylococci and mammaliicocci (NASM) cause mastitis and may contaminate milk and dairy products. Milk samples (n&#xa0;=&#xa0;1916) from cows with subclinical or clinical mastitis (SCM and CM, respectively) were collected from 28 large-scale (> 500 lactating cows) Chinese dairy farms. Overall, 999 NASM isolates representing 19 species were identified by MALDI-TOF MS and cpn60 sequencing, with Staphylococcuschromogenes, Mammaliicoccus sciuri and Staphylococcus haemolyticus being most prevalent. Antimicrobial resistance (AMR) was determined with disc diffusion; non-susceptible to penicillin was most common (SCM, 30% and CM, 29%) whereas cefoxitin non-susceptible NASM accounted for 8-10% of isolates; among these, 12.5% carried mecA but none carried mecC. Galleria mellonella was used to assess virulence of 78 strains of S. chromogenes, a dominant species; subsequently, 32 strains, representing higher- and lower-virulence in the Galleria model, were selected for whole-genome sequencing and comparative genomics. S. chromogenes isolates from CM had higher virulence (p&#xa0;<&#xa0;0.05) than those from SCM. The 32 genomes comprised 20 sequence types, indicating high genetic diversity. No robust genomic marker of Galleria virulence phenotype was identified in this selected WGS subset. Acquired resistance genes (n&#xa0;=&#xa0;5) were detected, including a first report of fusC in S. chromogenes; the fusC-positive isolate had an elevated fusidic acid MIC (8&#xa0;mg/L). Although S. chromogenes persisted in milk at 4&#xa0;&#xb0;C, pasteurization (64&#xa0;&#xb0;C for 30&#xa0;min) reduced viable counts to below detection. This study provided new insights into the prevalence, AMR, genomic diversity, and dairy-chain relevance of milk-derived NASM, particularly S. chromogenes. However, the genomic findings were based on an intentionally selected WGS subset and should be interpreted as hypothesis-generating rather than population-representative.

Animals