Search PubMedSearch

SEARCH · Search PubMed

Results for “Sequence Deletion”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

332 records · Page 4Linked to original sources

A genome-wide coverage-based pipeline for the identification of host-derived candidate DNA biomarkers from cell-free blood.

We have created a new data-analysis pipeline for the discovery of host-specific candidate DNA biomarkers derived from sequencing data of cell-free blood. Unlike approaches that rely on specific molecular or genetic signatures, our method leverages the coverage distribution of cell-free DNA sequences mapped to a reference genome, applying statistical analyses to identify informative short genomic regions for biomarker discovery. The pipeline is applicable to diverse diseases and can be used to analyze cell-free DNA sequences from plasma or serum to identify candidate biomarkers that are characteristic of disease states in mammals. Core functionalities were developed in Java and integrated with open-source software tools for the preprocessing of raw sequencing data, complemented by Python scripts for the machine-learning analysis and statistical validation. The pipeline is designed for HPC use and users can access the pipeline through a Galaxy workflow, which offers a user-friendly web interface for input selection prior to execution and analysis progress monitoring. Performance tests, carried out using duplicate sets of COVID-19 samples and controls, showed linear scalability of execution time with an increasing dataset size, as well as a substantial reduction in execution time through parallelized computation, whereby each HPC node is used to process the data of one chromosome. Further statistical tests confirmed the quality of the pipeline's results by showing that the set of identified candidate biomarkers remained stable across varying dataset sizes.

Biomarkers

Comparison of paralog identification methods and their impact on species tree topologies in target capture phylogenomics within the Sindora clade (Detarioideae: Leguminosae).

Target capture is a common method of generating high throughput DNA sequencing data for phylogenetic reconstruction of species relationships, for which single copy genes are usually most informative. However, a pervasive problem with target capture is that putatively single copy genes may in fact be paralogs resulting from gene duplication, which are problematic for phylogenetic inference because their evolutionary history may differ from the divergence history of species. Here, we use as a case study a target enrichment dataset of 88 species of Detarioideae (Leguminosae) with a focus on the Sindora clade to examine approaches for handling paralogs, including the built-in paralog handling functions in HybPiper and CAPTUS, plus subsequent steps using Putative Paralog Detection and the tree-based Yang & Smith orthology inference approach. We compare the paralogs flagged using these methods and verify their performance with BLAST mapping against a reference genome sequence of Sindora glabra, and then subsequently compare the species tree topologies produced across these methods. Our comparisons of paralogs flagged across the Sindora clade show that the Putative Paralog Detection pipeline was the most accurate in identifying paralogs in terms of its similarity to the BLAST mapping, followed by the built-in paralog identification function of CAPTUS. However, the results we recovered for the Detarioideae subfamily suggest that the largest differences in species tree topology resulted from the use of paralog-filtered alignments (such as with the Putative Paralog Detection pipeline and the Yang & Smith orthology inference approaches) rather than just by removing the sequences of identified paralogous genes. This was the true for HybPiper-assembled datasets but was not seen in CAPTUS-assembled datasets. In all comparisons, the topological differences caused by different paralog handling methods tended to be confined to clades where processes such as hybridisation and introgression are prevalent. Our study provides a roadmap to establish the best approach to identify, eliminate or separate paralogs in the absence of a chromosomally contiguous reference genome for a study group, and highlights the importance of careful data inspection and processing in addition to understanding the extent of paralogy and paralog characteristics (e.g. sequence divergence between copies) for their study group.

Phylogeny

Conserved host-exclusive oligonucleotide motifs enriched in pathogenic genes of human oncogenic viruses.

Comparative viral genomics can reveal sequence-level constraints influencing virus-host interactions. Relative minimal absent words (rMAWs) are short oligonucleotide motifs present in viral genomes but completely absent from the host, potentially reflecting selective pressures related to host adaptation and immune evasion. Using the EAGLE algorithm and the GRCh38 human reference genome, we systematically screened for prevalent rMAWs (prMAWs) across six major human oncogenic viruses: Epstein-Barr virus (EBV), hepatitis B virus (HBV), hepatitis C virus (HCV), human papillomavirus (HPV), human T-cell leukemia virus type 1 (HTLV-1), and human herpesvirus 8/Kaposi's sarcoma-associated herpesvirus (HHV-8/KSHV). highly conserved 11- and 12-bp prMAWs were identified in EBV, HBV, HTLV-1, and HHV-8/KSHV, with sequence prevalences ranging from 91.5% to 97.9%. Conversely, no short prMAWs were detected in HCV or HPV, likely reflecting differences in genome architecture, mutation rates, and long-term host adaptation to the human host. Importantly, the identified host-exclusive motifs exhibited non-random genomic distribution and were preferentially embedded within viral genes central to replication, persistence, immune modulation, and oncogenesis, including EBNA-1 (EBV), HBx (HBV), Tax-associated regions (HTLV-1), and lytic replication genes of HHV-8/KSHV. Notably, all detected prMAWs were enriched in GC nucleotides and exhibited marked CpG over-representation, suggesting sequence constraints associated with epigenetic regulation and viral persistence. Collectively, these highly conserved, host-exclusive signatures offer promising, candidates for sequence-directed approaches in the diagnosis, monitoring, and investigation of virus-associated cancers.

Humans

A conserved distal-tail helical extension defines a tailspike attachment architecture in Gram-negative siphophages.

Rapid growth of bacteriophage genome collections has outpaced functional annotation of tail-tip proteins, limiting comparative analysis of host-recognition structures. Starting from a shared distal-tail gene organization in the Salmonella phages 9NA and Jersey, I developed a morphogenetic bioinformatic framework integrating gene synteny, sequence comparison, profile hidden Markov model (HMM) screening, structural evidence, structure-aware searching, and AlphaFold modeling. Comparison with the experimentally characterized lambda and Sf11 tail assemblies identified a predominantly alpha-helical C-terminal extension of the distal-tail (DT) protein associated with tailspike attachment, termed the distal-tail helical extension (DT-helix). Screening 541,986 proteins from 5167 complete NCBI RefSeq tailed-phage genomes, followed by evidence-based evaluation of sequence, genomic context, and structural architecture, identified 165 curated DT-helical-extension-associated phages. Their DT proteins segregated into six sequence groups. In the four principal multi-member groups, cognate tailspikes showed group-specific conservation in proximal N-terminal regions but substantially greater downstream diversity, consistent with sequence constraint at the DT-tailspike attachment boundary. A complementary ProstT5/Foldseek search supported the established groups but revealed no convincing additional highly divergent family. Together with the experimentally characterized Sf11 attachment interface, these findings define a recurrent morphogenetic architecture linking conserved distal-tail scaffolds to more variable receptor-binding proteins across siphophages infecting Gram-negative bacteria. Although universal exchangeability is not established, the identified scaffold-receptor-binding boundaries provide a framework for molecular characterization and rational phage engineering. Accession-level information for the 165 curated phages is available through PhageTailDB.

Viral Tail Proteins

Mitochondrial DNA diversity in Ecuadorian populations: Recurrence of variant 16136 within haplogroup B2.

The identification of lineage-defining variants, frequently found in the coding region of mitochondrial DNA (mtDNA), is essential for refining haplogroup classification. Most mtDNA studies in South American populations have focused on the control region (CR), which has provided important insights into population structure and maternal lineage origins, although information needed for more robust phylogenetic resolution has been neglected. This study investigates the maternal genetic structure of Ecuadorian populations by combining CR and whole mitogenome analyses. Sequences from the mtDNA CR were obtained from 461 individuals (253 Mestizos and 208 Native Americans), while complete mitogenomes were sequenced for 127 individuals to improve phylogenetic resolution by identifying lineage-defining variants present in coding region. Most mtDNA haplogroups in the two population groups analyzed were of Native American origin (A2, B2, B4, C1, D1, D4), with significant differences in the distribution of specific lineages between them. Among Mestizos, African haplogroups (all within the L branches) and Eurasian haplogroups (H, K, R, U) were detected at low frequencies, whereas no African lineages were observed among Native Americans. The results obtained highlighted a heterogeneity within Ecuadorian populations that must be considered when developing mtDNA haplotype databases for forensic purposes. Whole mitogenome sequences enabled the identification of variants that refined haplogroup classifications, provided a more accurate reconstruction of the maternal genetic diversity, and improve the discrimination between Native American and Asian maternal lineages within haplogroup B4b.

Humans

Real-world clinical utility of exome sequencing in pediatric drug-resistant epilepsy: Experience from a tertiary center in Thailand.

BACKGROUND: Genomic testing has increasingly contributed to the diagnosis and management of pediatric drug-resistant epilepsy (DRE), particularly in patients with suspected genetic etiologies. This study evaluated the diagnostic yield and real- world clinical utility of whole-exome sequencing (WES) in children with DRE. METHODS: Children with DRE and seizure onset before 15 years of age were enrolled between January 2020 and December 2023. Clinical data, including demographics, seizure characteristics, developmental history, electroencephalography (EEG), brain magnetic resonance imaging (MRI), and prior investigations, were reviewed. WES was performed in all probands and, when available, their parents. Variants were interpreted according to standard guidelines. Clinical utility and 1-year seizure and developmental outcomes were assessed from follow-up records. RESULTS: Fifty-six patients (23 males, 33 females) were included. The median age at seizure onset was 1 year (interquartile range [IQR] 0.3-4 years), and 96.4% had developmental comorbidities. Pathogenic or likely pathogenic variants were identified in 39% (22/56), with the highest diagnostic yield in children with seizure onset before 3 years of age. Channelopathies accounted for most genetically solved cases (68%), predominantly involving sodium channel genes. Genetic diagnoses provided clinical utility in 73% (16/22) of solved cases by guiding treatment and precision management. At 1-year follow-up, genetically solved patients showed more favorable seizure and developmental outcomes than those with genetically unsolved patients. CONCLUSION: WES achieved a 39% diagnostic yield and substantial clinical utility in pediatric DRE, particularly in early-onset and channelopathy-related disorders. These findings support early molecular diagnosis to facilitate genotype-informed management in appropriately selected children. However, the more favorable developmental and seizure outcomes observed in genetically solved patients should be interpreted with caution, as they may have been influenced by multiple factors beyond genetic diagnosis. In resource-limited settings, careful clinical phenotyping remains essential for treatment decisions and for prioritizing children for genomic testing.

Clinical utility

Characteristics of p53 and Smad4 immunohistochemistry in pancreatic ductal adenocarcinoma and validation by next-generation sequencing.

BACKGROUND: Mutations in four major driver genes -KRAS, CDKN2A, TP53, and SMAD4- are central to the pathogenesis of pancreatic ductal adenocarcinoma (PDAC) and critically inform diagnosis, therapeutic decision-making, and prognostic assessment. Although next-generation sequencing (NGS) is widely regarded as the gold standard for detecting these mutations, its clinical application is often limited by suboptimal analytical efficiency and substantial economic cost. Among these genes, immunohistochemical (IHC) staining for the proteins encoded by TP53 and SMAD4 has been extensively adopted in routine pathology practice. However, standardized IHC pattern classification schemes and rigorous validation of their predictive accuracy for underlying genomic alterations remain lacking in PDAC. METHODS: We retrospectively enrolled 63 PDAC patients and systematically characterized the typical IHC expression patterns of p53 and Smad4. Targeted NGS was subsequently performed on all available tumor specimens, and the resulting mutational profiles were correlated with corresponding IHC findings. Diagnostic performance including sensitivity, specificity and accuracy of p53 IHC for predicting TP53 mutations and of Smad4 IHC for predicting SMAD4 mutations was rigorously evaluated. RESULTS: Among the four canonical driver genes, co-occurring double- or triple-gene mutations were prevalent; within TP53 and SMAD4, missense mutations constituted the most frequent variant type. Using NGS as the reference standard, we validated the diagnostic utility of a three-tiered p53 IHC classification system, particularly in fine-needle biopsy (FNB) specimens. Furthermore, we proposed a novel, refined Smad4 IHC pattern classification that incorporates an "intermediate" category, thereby expanding upon conventional binary interpretation. This new scheme achieved markedly improved mutation prediction accuracy (0.76) compared with traditional approaches (0.57). CONCLUSION: Our study highlights the complementary diagnostic value of p53 and Smad4 IHC relative to molecular testing in PDAC, especially when tissue is limited, as commonly encountered in FNB specimens. The newly established Smad4 IHC classification system, which integrates an intermediate expression category into the conventional two-tier framework, demonstrates superior clinical utility and enhances predictive accuracy for SMAD4 genomic alterations.

Humans

Longitudinal whole-genome analysis of bluetongue virus identifies conserved serotype-specific genomes and distinct genomic constellations within a Colorado sheep flock (2021-2023).

Bluetongue virus (BTV) is a segmented double-stranded RNA virus of ruminants transmitted by Culicoides spp. biting midges. Although the genome consists of ten segments, classification into serotypes is primarily based on genome segment 2. However, reassortment among genomic segments is a major driver of BTV evolution and diversity. This study used longitudinal whole-genome sequencing to characterize BTV genomes collected from 2021 to 2023 within a single sheep flock in Colorado, where multiple serotypes co-circulate. Whole-genome sequences were generated from fourteen blood samples representing four serotypes: BTV-6, -11, -13, and -17. Longitudinal sampling identified multiple BTV serotypes within individual sheep across consecutive years. Tanglegram analysis comparing segment phylogenies to the segment 2 tree demonstrated incongruent topologies across all genomic segments, suggestive of reassortment or the circulation of distinct genomic constellations. Nucleotide-level comparisons revealed high sequence homology among same-serotype samples from the same year, while the greatest genetic divergence was observed among BTV-17 genomes collected in different years. Additionally, all BTV-13 genomes contained a previously undescribed nonsynonymous substitution in segment 10 predicted to extend the encoded protein by three amino acids. Together, these findings demonstrate that highly conserved BTV genomes and distinct genomic constellations can be detected at the flock level across multiple years. This longitudinal whole-genome approach reveals the genetic complexity of endemic BTV populations, including novel variants and genomic patterns consistent with reassortment that are lost with conventional serotyped-based approaches, highlighting the need to integrate whole-genome characterization into endemic BTV monitoring programs.

Animals

Trio-based whole-exome sequencing identifies convergent epithelial junction-related pathways in syndromic hidradenitis suppurativa.

INTRODUCTION: Hidradenitis suppurativa (HS)-related autoinflammatory syndromes, simply termed as syndromic HS (sHS), represent a group of rare immune-mediated inflammatory disorders in which HS coexists with systemic or cutaneous autoinflammatory features like PASH (pyoderma gangrenosum-PG-, acne and HS), PAPASH (PASH, pyogenic arthritis), PASS (PG, acne, HS, and ankylosing spondylitis), and SAPHO syndrome (synovitis, acne, pustulosis, hyperostosis, and osteitis). In recent years, genetic studies identified several novel pathogenic variants underlying sHS; however, most investigations rely exclusively on affected individuals sequencing and the absence of parental genomic information limits the possibility to determine inheritance patterns. METHODS: To address these gaps, we performed trio-based whole-exome sequencing (WES) on five individuals diagnosed with sHS and their unaffected parents. RESULTS: The pathway related to epidermal adhesion and desmosome organization was the most represented across our cohort, encompassing seven genes: DSC3, DSG1, FAT1, LAMA3, MICALL2, PLEC and TJP2. Integrin-extracellular matrix (ECM) adhesion signaling pathway, represented by ten genes (CSPG4, FERMT3, ITGA3, LAMA3, LAMA5, LIMS2, LTBP3, PLEC, TGM2, TNC) was also retrieved. Also, variants affecting innate immune pathways, including cytokine signalling and antigen presentation, have been observed. CONCLUSION: Our exploratory findings suggest that genetically heterogeneous variants in syndromic HS converge on biological processes involving epithelial junction organisation, extracellular matrix interactions and innate immune regulation. Although not establishing a unique pathogenic mechanism, these observations identify epithelial barrier biology as a candidate pathway warranting validation in larger cohorts and functional studies.

Journal Article

Functions of tandem-repeat galectins and domain coordination governs galectin-4 activity in grass carp (Ctenopharyngodon idella).

Galectins are β-galactoside-binding lectins that play essential roles in innate immunity. Among them, tandem-repeat galectins (TrGals), typically composed of two distinct carbohydrate-recognition domains (CRDs) connected by a linker peptide, are well established as key regulators of pathogen recognition and host defense in mammals. However, their structural diversity and immunological functions in teleost fish remain poorly understood. In this study, five TrGals (Gal-4, Gal-8a, Gal-8b, Gal-9, and Gal-9like) were identified in grass carp. Sequence and structural analysis revealed that Gal-8a/b, Gal-9, and Gal-9like possess the canonical two-CRD architecture, whereas Gal-4 uniquely contains four highly similar tandem-repeat domains. All five TrGals were broadly expressed across examined tissues, with predominant expression in the liver. Upon Aeromonas hydrophila infection, Gal-4, Gal-8a, Gal-8b, and Gal-9 were rapidly up-regulated at early time points (3-6 h). To elucidate the functional significance of CRD number, recombinant full-length CiGal-4 (CiGal4-full) and three truncated variants containing one, two, or three CRDs (CiGal4-1CRD, CiGal4-2CRD, and CiGal4-3CRD) were generated and systematically characterized. All recombinant proteins contained the conserved β-sheet structure typical of galectin CRDs. Functional assays revealed that CiGal4-full displayed the strongest growth-inhibitory activity against all tested bacteria, whereas CiGal4-1CRD showed the weakest effect. Notably, CiGal4-2CRD exhibited the most potent bactericidal activity, surpassing the full-length protein, while CiGal4-3CRD showed no further enhancement. CiGal4-full and CiGal4-2CRD showed superior carbohydrate-binding activities compared with the other variants. Collectively, these results reveal that CRD copy number alone does not linearly determine galectin function. Instead, domain organization and conformational coordination are critical for optimizing antimicrobial activity. This study provides new insights into the structure-function relationships and evolutionary diversification of galectins in teleosts and highlights their potential as novel antimicrobial and immunomodulatory agents in aquaculture.

Animals

Discovery and characterization of multifunctional bioactive peptides from Alaska Pollock (Gadus chalcogrammus) milt: hybrid in silico, in vitro, and proteomic approaches.

The growing demand for multifunctional bioactive peptides has sparked interest in underutilized marine by-products as sustainable bioresources. This study explored Alaska Pollock (Gadus chalcogrammus) milt protein as a novel source of peptides with anti-inflammatory, anti-hypertensive, and anti-diabetic effects. Protein composition was analyzed via LC-MS, followed by in silico digestion and bioactivity prediction. Molecular docking identified peptides targeting DPP-IV, α-glucosidase, ACE, GLP-1 receptor, COX-2, MuRF1, and the 20S proteasome. Among the candidates, a promising peptide (CLPPH) was synthesized and validated in vitro, demonstrating inhibitory effects on nitric oxide production, DPP-IV, ACE, and α-glucosidase. These results highlight CLPPH's potential as a multifunctional bioactive peptide and support the valorization of Alaska Pollock milt as a sustainable source for functional foods and nutraceutical applications.

Animals

Gene-environment interaction between perinatal oxytocin exposure and Pten mutation shapes epigenetic reprogramming of oxytocin signaling and behavior in mice.

Synthetic oxytocin (Pitocin) is the most commonly used pharmacologic agent for induction and augmentation of labor. Beyond its uterotonic effects, oxytocin plays a critical role in neurodevelopment and social behavior. Dysregulated oxytocin signaling has been implicated in autism spectrum disorder (ASD), raising concern that perinatal exposure to exogenous oxytocin may have lasting neurodevelopmental consequences. This study aimed to determine whether offspring harboring a genetic predisposition for ASD are differentially impacted by perinatal oxytocin exposures, with a focus on long-term oxytocin signaling and autism-like behavior. Pregnant mice carrying offspring with heterozygous mutations in phosphatase and tensin homolog deleted on chromosome ten (Pten), a well-established monogenic risk factor for ASD, received continuous oxytocin versus phosphate-buffered saline (PBS) control via micro-osmotic pumps during late gestation. Wild-type (WT) offspring exposed to each treatment served as a secondary control. Adult offspring were assessed for oxytocin receptor (Oxtr) methylation in the frontal cortex and hippocampus, oxytocin expression in the hypothalamus, serum oxytocin levels, and were subject to a battery of social and anxiety-related behavior tests. Perinatal oxytocin exposure produced genotype-dependent effects in offspring. Epigenetic analyses revealed bidirectional remodeling of Oxtr methylation in the frontal cortex and hippocampus, with increased exon 1 methylation in WT mice and decreased methylation in Pten-mutant mice, resulting in significant genotype-treatment interactions. Hypothalamic oxytocin expression increased following treatment regardless of genotype, though baseline levels were higher in Pten-mutant mice. Neither oxytocin treatment nor genotype impacted long-term serum oxytocin levels. Behavioral outcomes were modest but context-specific: repetitive behaviors and cognition performance were unchanged, but oxytocin-treated Pten-mutant mice exhibited increased anxiety-like behavior alongside improved social memory. In contrast, oxytocin-treated WT mice showed reduced social novelty preference. Exploratory analyses suggested potential sex-dependent trends. Our findings support a model in which genetic susceptibility shapes the epigenetic encoding of early-life hormonal signals, thereby recalibrating oxytocin system function and downstream behavioral outcomes. Together, these data highlight the context-dependent effects of perinatal oxytocin exposure and argue against uniformly beneficial or detrimental effects, emphasizing the importance of gene-environment interactions in neurodevelopmental trajectories.

Animals

Genomic characterization and pathogenicity of ruminant Listeria monocytogenes isolates in a murine oral infection model.

Listeria monocytogenes is a major foodborne pathogen; its ruminant isolates display zoonotic characteristics, causing similar clinical signs in humans, including abortion and encephalitis. However, data on whole genome sequencing and pathogenicity of ruminant L. monocytogenes isolates remain sparse. This study aimed to analyze the genotypic characteristics of L. monocytogenes isolates from ruminants with listeriosis. Furthermore, we assessed the in vivo pathogenicity of four ruminant L. monocytogenes isolates, characterized via whole-genome sequencing-based genetic clustering, in orogastrically inoculated mice. The isolate LM18 (serotype 1/2b, ST224, SL6178) had the lowest lethal dose compared to the other three isolates including previous hypervirulence type (serotype 4b, ST1, SL1) and caused secondary bacteremia in lungs, with sustained bacterial loads in the spleen and liver. Genomic (listeria pathogenicity island -1 and -3) and virulence gene (actA and llsX) mutation analyses associated with virulence suggested from well-recognized studies could not elucidate the virulence of the isolates. SSI-1, which only exists in the isolate LM18 (serotype 1/2b, ST224, SL6178), may help L. monocytogenes survive in the gastrointestinal environment, thereby affecting its virulence. Further research should investigate the role of SSI-1 in the pathogenicity of L. monocytogenes. Moreover, additional studies utilizing larger datasets of ruminant isolates are required to validate our genotypic characterization and to obtain a comprehensive picture of further genotypic differences crucial for L. monocytogenes pathogenicity.

Animals

Exploring the mechanism of aroma production in fermented cherry juice by L. brevis LD1.0600 using flavomics and whole genome analysis.

This study focused on L.brevis LD1.0600 with excellent fermentation traits: it analyzed genome-wide key regulatory genes for micro-metabolites, combined with fermented cherry juice flavor metabolomics data, and used machine learning to explore correlations between gene regulation, metabolite production, and flavor formation. The SVM model screened and verified fermented cherry juice VOCs; through OAV and flavor wheel analysis, LD1.0600 emerged as the top-performing strain, with a sweet, fruity dominant aroma. Key aroma-active components (OAV > 100) included 2-methoxy-4-vinylphenol, benzaldehyde, 2-methyl-butanoic acid and hexanoic acid, and 2-methoxy-4-vinylphenol and hexanoic acid elevated by LD1.0600-regulated genes (Chrom1-001884, Chrom1-000925, fabF and Chrom1-000199). At the same time, through research, a "strain screening-SVM screening of DVCs-OAV screening of key aroma components-whole genome sequencing of flavor regulatory genes" system was established. This system can not only be applied to the screen fermentation strains, but also can be extended to the application of other fermentation products.

Fermentation

Translating single-cell RNA sequencing into monocyte direct leukocyte subpopulation-transcript abundance assay ratio-based biomarkers (IFI27/PSAP or IFI27/CTSS) for clinical detection of viral infection.

A rapid method for triaging febrile patients by aetiology (e.g., viral or bacterial infection) using gene expression in peripheral blood (PB) is an intensively researched area. However, gene expression in blood represents a composite sum of gene expression of all the component cell types present in the sample. As a result, numerous genes are measured in most proposed signatures. Herein, we propose a simple ratio-based biomarker (RBB) called direct leukocyte subpopulation-transcript abundance assay (DIRECT LS-TA) that recapitulates gene expressions of a single cell type in PB (i.e., monocytes). Based on single-cell RNA sequencing (scRNAseq) data and bulk expression data, IFI27 and SIGLEC1 are found as interferon-stimulated genes (ISGs) predominantly expressed by monocytes. The DIRECT LS-TA method can use a simple ratio of two genes measured in PB as an RBB to represent the target gene expression in monocytes without the need for monocyte purification. Both scRNAseq and bulk RNA sequencing datasets were used to evaluate the correlation between ISG expression in monocytes and PB, with a particular focus on monocyte expression of IFI27. An iceberg plot of bulk transcriptome data was used to identify genes that were predominantly expressed by monocytes in PB. DIRECT LS-TA RBBs of the three genes (IFI27, IFI44L and SIGLEC1) were evaluated by group-wise comparison, receiver operating characteristic and meta-analysis. In addition, the conventional interferon (IFN) score was evaluated for comparison of diagnostic performance. In viral infection datasets, DIRECT LS-TA of IFI27 (IFI27/PSAP or IFI27/CTSS) was most intensely activated (p value by t test <1e-9) and had the best area under the curve (0.94) among the three potential monocyte ISGs analysed. DIRECT LS-TA SIGLEC1 was also another monocyte biomarker but showed a lower activation (p<9e-5). IFI27/PSAP showed better diagnostic performance than the conventional IFN score. On the other hand, IFI44L was not a predominant monocyte expression gene. DIRECT LS-TA of IFI27 (IFI27/PSAP or IFI27/CTSS) measured in PB was the best biomarker of viral infection and IFN activation among ISGs predominantly expressed by monocytes. It performed even better than the conventional IFN score which required quantification of eight genes. The results suggest that DIRECT LS-TA of IFI27 is a monocyte-informative biomarker which is easy to determine in PB without the need for cell sorting.

Humans

Insights into the fate and dynamics of antibiotic resistance in multidrug-resistant Bacillus cereus during in vitro simulated gastrointestinal digestion.

Bacillus cereus, an important pathogen responsible for causing foodborne diseases worldwide, releases pore-forming enterotoxins, which target host epithelial cells, leading to osmotic lysis and ultimately manifesting as diarrheal syndrome. Moreover, some B. cereus strains carry antimicrobial resistance genes that confer multidrug resistance against a spectrum of antibiotics. Characterizing the survival traits of multidrug-resistant (MDR) B. cereus strains in the intestinal microenvironment is essential for developing targeted strategies to effectively manage diarrheal foodborne diseases caused by this pathogen. This study used whole-genome sequencing (WGS) to evaluate the pre- and post-digestion toxigenic potential, antimicrobial resistance profiles, and genetic diversity of MDR B. cereus strains isolated from food samples in Guangdong Province, China. The four B. cereus isolates investigated in this study exhibited a genetic diversity, as determined by multilocus sequence typing analysis of WGS data. All four isolates produced the diarrheal toxins Hbl, Nhe, and CytK to varying levels, indicative of their potential to cause outbreaks of foodborne diseases. Each of the four isolates exhibited resistance to more than three classes of antibiotics, fulfilling the criterion for multidrug resistance. At an initial concentration of 9 log colony-forming units (CFU)/mL, the intestinal concentration of these four isolates crossed the threshold required to induce widespread diarrhea in the general population. Under rice slurry protection, all tested isolates maintained intestinal concentration beyond the threshold when the initial concentration was increased to &#x2265;8 log CFU/mL. Moreover, the upregulations of genes associated with acid tolerance, bile tolerance and stress response were observed in the surviving MDR B. cereus isolates. Digestion markedly altered the antibiotic resistance profiles of the MDR B. cereus isolates. In the absence of a food matrix, the MDR isolates lost their resistance to imipenem, meropenem, amoxicillin-clavulanic acid, and trimethoprim-sulfamethoxazole post-digestion and was influenced by the initial concentration of the strains. In the presence of food matrix rice slurry, the effects of digestion on the antibiotic resistance of MDR B. cereus isolates can be mitigated, enabling them to maintain their antibiotic resistance to the greatest extent. Most remarkably, after digestion, the isolates Bce055 and Bce166 exhibited newly emergent resistance to cefotetan and trimethoprim-sulfamethoxazole, respectively. Our findings clarify the fate of MDR B. cereus isolates in the gastrointestinal tract and inform the development of prevention and control strategies for foodborne diseases caused by this pathogen.

Drug Resistance, Multiple, Bacterial

Comparative genomic epidemiology of food- and patient-derived diarrheagenic Escherichia coli from sentinel surveillance in Southeast China.

Diarrheagenic Escherichia coli (DEC) remains an important foodborne pathogen, yet long-term comparative genomic surveillance data jointly characterizing food-derived and patient-derived isolates remain limited. This surveillance-based comparative study integrated antimicrobial susceptibility testing and whole-genome sequencing to characterize diarrheagenic Escherichia coli isolates recovered from food and patient sources in Lishui, Southeast China, during 2018-2025, with emphasis on occurrence, resistance profiles, genomic backgrounds, and plasmid replicon-associated features. Antimicrobial susceptibility testing was performed for 258 selected isolates, and whole-genome sequencing was conducted for a curated analytical subset of 204 isolates. The sequenced subset was used for diversity-oriented comparative genomic analysis rather than for unbiased prevalence estimation of the entire DEC collection. EAEC predominated in both sources, although food-associated occurrence was heterogeneous across categories, with the highest recovery rate observed in raw meat. Patient-derived isolates showed a broader overall resistance burden, whereas food-derived isolates retained substantial resistance to tetracycline, chloramphenicol, and florfenicol. Phylogenetic analysis showed partial overlap in genomic backgrounds between food-derived and patient-derived isolates, while representative resistance determinants displayed both broadly distributed and lineage-enriched patterns. Replicon-based plasmid profiling identified 42 plasmid types, including 12 detected in both sources, with IncF-related replicons predominating among these shared profiles. Several food-derived isolates carried multiple plasmid replicon types that were also observed in patient-derived isolates. Overall, food-derived and patient-derived DEC showed partial overlap in genomic backgrounds, resistance determinants, and replicon-defined plasmid profiles within this surveillance setting, while retaining source-associated heterogeneity. These findings should be interpreted as surveillance-based comparative evidence rather than as evidence of direct source attribution or transmission.

Humans