Search PubMedSearch

SEARCH · Search PubMed

Results for “Molecular Sequence Data”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

993 recordsLinked to original sources

Discovery and characterization of multifunctional bioactive peptides from Alaska Pollock (Gadus chalcogrammus) milt: hybrid in silico, in vitro, and proteomic approaches.

The growing demand for multifunctional bioactive peptides has sparked interest in underutilized marine by-products as sustainable bioresources. This study explored Alaska Pollock (Gadus chalcogrammus) milt protein as a novel source of peptides with anti-inflammatory, anti-hypertensive, and anti-diabetic effects. Protein composition was analyzed via LC-MS, followed by in silico digestion and bioactivity prediction. Molecular docking identified peptides targeting DPP-IV, α-glucosidase, ACE, GLP-1 receptor, COX-2, MuRF1, and the 20S proteasome. Among the candidates, a promising peptide (CLPPH) was synthesized and validated in vitro, demonstrating inhibitory effects on nitric oxide production, DPP-IV, ACE, and α-glucosidase. These results highlight CLPPH's potential as a multifunctional bioactive peptide and support the valorization of Alaska Pollock milt as a sustainable source for functional foods and nutraceutical applications.

Animals

Upscaling Genotyping by Amplicon Sequencing With GBAS-GUI.

Genotyping by amplicon sequencing (GBAS) is a relatively low-cost approach for generating genotypic data compared with established genomic methods, making it highly scalable and particularly suitable for large-scale genetic monitoring projects. However, most existing analytical pipelines are either marker-specific, insufficiently scalable, or lacking efficient data management systems for the long-term integration of genotypic information, limiting the full potential of GBAS. Here, we address this gap by introducing GBAS-GUI (https://github.com/sonnenbe-dot/GBAS-GUI), a pipeline capable of generating GBAS-based genotypic data for a wide variety of loci at scale. GBAS-GUI integrates a graphical user interface with multiple checkpoints to improve accessibility and robustness. It implements multiprocessing architecture and a relational database that links genotypic data with associated sample metadata to enhance scalability and data management. The pipeline further enables marker screening through automated calculation of polymorphism information content (PIC) and implements a strategy to recover homologous genotypic information from paralogous loci with non-overlapping amplicon length ranges. Using multiple empirical datasets, we demonstrate substantial improvements in processing speed, database management and handling artefacts related to co-amplification of unspecific regions and duplicates of the same genomic region. We further show that incorporating the full sequence information captured by an amplicon increases marker information content beyond what is achievable with length-based genotyping alone and expands the analytical versatility of GBAS. Overall, GBAS-GUI provides a robust, scalable and versatile framework that unlocks the potential of GBAS for large-scale population genetic and phylogeographic studies.

Genotyping Techniques

Complex evolutionary history of Rosales mediated by extensive incomplete lineage sorting and hybridization.

The angiosperm order Rosales still represents a major challenge for phylogenetic reconstruction. Although its circumscription is now well-defined, phylogenetic relationships among families are still uncertain. Here, we used nuclear, plastid, and mitochondrial genomic data from 33 species representing all nine families to further clarify interfamilial relationships and the group's evolutionary history. We detected significant phylogenetic conflict among the three datasets. Further analyses at the nuclear level identified incomplete lineage sorting (ILS) as the main cause of unstable phylogenetic positions among families. The discordant placements of Rhamnaceae and Elaeagnaceae based on plastid and mitochondrial data are caused by ancient hybridization events, potentially involving differences in organellar inheritance. Our molecular dating confirms earlier suggestions that the ancient rapid diversification of the three Rosaceae subfamilies could be the main reason for the difficulties in resolving their phylogenetic relationships. Our findings provide new insights into the interfamilial relationships of Rosales and demonstrate that the evolutionary history of this order was shaped by ancient and rapid radiation as well as extensive ILS and reticulate evolution. They also suggest that previous attempts to clarify interfamilial relationships in this order were hampered by combining nuclear and organellar sequence data, leading to inconsistent topologies observed across earlier studies.

Phylogeny

First identification and molecular subtyping of Blastocystis spp. in donkeys in Aksaray province, Türkiye.

Blastocystis is a common intestinal protist worldwide that can infect humans and animals. Although its molecular epidemiology in Türkiye is mostly focused primarily on humans and livestock, equids have received limited attention despite their traditional roles and frequent contact with humans and other animals in rural environments. This study aimed to determine the molecular prevalence and subtype (ST) distribution of Blastocystis spp. in donkeys in Aksaray Province, providing the first molecular data on donkeys in Türkiye. A total of 182 fresh fecal samples were collected from donkeys in nine villages within Aksaray province. Genomic DNA was extracted, and the small subunit ribosomal RNA (SSU rRNA) gene fragment of Blastocystis spp. was amplified via PCR analysis. Positive isolates were sequenced bidirectionally for identification and subsequent phylogenetic analysis of Blastocystis in donkeys. The overall molecular prevalence of Blastocystis spp. in donkeys was 4.4% (8/182). The infection rate was higher in young donkeys (under 3 years old; 8.33%) than in adults (3 years or older; 2.46%). However, this difference was not statistically significant. Sequence analysis of the positive PCR products revealed the presence of one known livestock-specific subtype, ST10. Phylogenetic analysis showed that the ST10 isolates characterized in this study clustered with isolates identified from different hosts. This study provides the first molecular data on Blastocystis presence in donkeys in Türkiye. The exclusive detection of ST10 suggests potential cross-species transmission, likely facilitated by the traditional practice of co-housing donkeys with other animals in confined barns. These findings indicate that donkeys may contribute to Blastocystis transmission, underscoring the importance of a "One Health" approach in future epidemiological surveillance.

Animals

Comparison of paralog identification methods and their impact on species tree topologies in target capture phylogenomics within the Sindora clade (Detarioideae: Leguminosae).

Target capture is a common method of generating high throughput DNA sequencing data for phylogenetic reconstruction of species relationships, for which single copy genes are usually most informative. However, a pervasive problem with target capture is that putatively single copy genes may in fact be paralogs resulting from gene duplication, which are problematic for phylogenetic inference because their evolutionary history may differ from the divergence history of species. Here, we use as a case study a target enrichment dataset of 88 species of Detarioideae (Leguminosae) with a focus on the Sindora clade to examine approaches for handling paralogs, including the built-in paralog handling functions in HybPiper and CAPTUS, plus subsequent steps using Putative Paralog Detection and the tree-based Yang & Smith orthology inference approach. We compare the paralogs flagged using these methods and verify their performance with BLAST mapping against a reference genome sequence of Sindora glabra, and then subsequently compare the species tree topologies produced across these methods. Our comparisons of paralogs flagged across the Sindora clade show that the Putative Paralog Detection pipeline was the most accurate in identifying paralogs in terms of its similarity to the BLAST mapping, followed by the built-in paralog identification function of CAPTUS. However, the results we recovered for the Detarioideae subfamily suggest that the largest differences in species tree topology resulted from the use of paralog-filtered alignments (such as with the Putative Paralog Detection pipeline and the Yang & Smith orthology inference approaches) rather than just by removing the sequences of identified paralogous genes. This was the true for HybPiper-assembled datasets but was not seen in CAPTUS-assembled datasets. In all comparisons, the topological differences caused by different paralog handling methods tended to be confined to clades where processes such as hybridisation and introgression are prevalent. Our study provides a roadmap to establish the best approach to identify, eliminate or separate paralogs in the absence of a chromosomally contiguous reference genome for a study group, and highlights the importance of careful data inspection and processing in addition to understanding the extent of paralogy and paralog characteristics (e.g. sequence divergence between copies) for their study group.

Phylogeny

Genome-wide SNP data support species boundaries in sympatric Polylepis Ruiz & Pav. (Rosaceae) species from Bolivia and Ecuador.

Species delimitation in the South American genus Polylepis is notoriously challenging due to high morphological similarity and phenotypic plasticity, likely driven by hybridization and gene flow. Previous phylogenetic studies suggested that genetic structure aligns more strongly with geography than with taxonomy, questioning existing species concepts and hampering conservation efforts. We used double-digest RAD sequencing (ddRADseq) to generate genome-wide SNP data for 11 Polylepis species sampled across multiple localities in Bolivia and Ecuador. Population genetic analyses, phylogenetic inference, and network approaches were combined to assess whether genetic structure aligns more closely with taxonomy or geography. Morphologically defined species formed largely cohesive genetic lineages across regions, with species identity explaining substantially more genetic variation than locality. While localized admixture and reticulation were detected among closely related taxa, widespread species showed strong genetic cohesion and clear separation from congeners. Our results indicate that the sampled Polylepis species from Bolivia and Ecuador maintain distinct genetic identities despite localized signals consistent with gene flow. This genome-wide support for current taxonomy highlights Polylepis as a valuable model for studying speciation under gene flow and indicates that multiple geographic sampling will be essential in reconstructing a robust phylogeny of the genus, with important implications for conservation planning in Andean montane forests.

Bolivia

Pervasive hybridization and introgression in Diervilleae (Caprifoliaceae).

Diervilleae (Caprifoliaceae) is a horticulturally important lineage with striking floral diversity and a long history of interspecific crossing, suggesting reticulate evolution. We integrated nuclear SNPs and whole plastome data to reconstruct a phylogenomic backbone for the tribe and to identify hybrids, cultivated accessions, and introgression among lineages. Nuclear and plastid phylogenies consistently recover Weigela and Diervilla as reciprocally monophyletic and resolve four major lineages within Weigela, providing a reproducible framework for revising sectional limits and species boundaries. Cultivated accessions form a well supported clade sister to W. florida and show predominantly W. florida ancestry while retaining contributions from multiple wild lineages, consistent with recurrent crossing, backcrossing, and selection. Analyses of wild populations reveal recurrent hybrids and enable plausible parental combinations to be inferred. Tests across the genome further indicate strong evidence for historical introgression across Diervilleae, with the strongest signals involving W. middendorffiana, W. maximowiczii, and Diervilla. Fossil evidence, divergence time estimation, and paleodistribution modelling together suggest range expansion during the Miocene and Pliocene followed by climate driven contraction, providing a spatiotemporal context for episodic contact, introgression, and the East Asia-North America disjunction.

Hybridization, Genetic

Genome-wide insights into the evolutionary and demographic history of the red alga Mazzaella laminarioides: Evidence for speciation with ancient migration along the southeast Pacific coast.

The mechanisms driving lineage divergence in red algae remain unexplored, despite the group's remarkable diversity and ancient evolutionary history. The red alga Mazzaella laminarioides, a Chilean intertidal species complex composed of three parapatric cryptic lineages (North, Center, South), offers a valuable system to evaluate these processes, as its life history combines severe dispersal limitation with a haploid-diploid cycle that may influence the emergence of reproductive barriers. We reconstructed its evolutionary history using whole-genome sequencing and nuclear genome assembly of representative individuals from each lineage. Phylogenomic analyses based on 1,507 single-copy orthologs recovered three deeply divergent lineages with limited nuclear discordance consistent with incomplete lineage sorting. For both splits, demographic modelling was most consistent with an Ancient Migration scenario, although support over strict isolation was moderate, suggesting that divergence may have begun with low asymmetric ancestral gene flow followed by subsequent loss of connectivity, demographic bottlenecks, and later population expansion. Coding sequence analyses revealed lineage-specific dN/dS heterogeneity; only one South-lineage locus passed FDR correction (metaxin-1, mitochondrial protein import), with two further South-lineage candidates in chlorophyll and heme biosynthesis falling below the FDR threshold. Together, these signals suggest that divergent selective pressures on energy acquisition may have contributed to divergence at the southern end of the distribution. These results add to the small but growing body of whole-genome data for red algae and, alongside recent macroalgal studies, suggest that ancestral connectivity could be a recurrent feature of lineage divergence even in marine organisms with extremely restricted dispersal.

Rhodophyta

Molecular Landscape and Advanced Diagnostic Technologies for BRAF Mutations in Cancer: From Quantitative PCR and ddPCR to CRISPR-Based Platforms.

BRAF mutations are key oncogenic alterations across multiple malignancies, including melanoma, thyroid carcinoma, colorectal cancer, non-small cell lung cancer, glioma, and hairy cell leukemia. The most prevalent variant, BRAF-V600E, induces constitutive activation of the MAPK signaling pathway, promoting tumor progression and influencing therapeutic responsiveness. Accurate detection of BRAF alterations is therefore essential for molecular classification, prognostic assessment, treatment selection, and resistance surveillance. This review summarizes the molecular heterogeneity of BRAF mutations and critically evaluates current diagnostic methodologies. Conventional approaches such as allele-specific PCR and Sanger sequencing are compared with advanced quantitative platforms, including high-resolution melting analysis, droplet digital PCR, and next-generation sequencing, with emphasis on analytical sensitivity, mutation coverage, and clinical applicability. Emerging technologies such as CRISPR-based assays, rolling circle amplification systems, and nanoparticle-based biosensors and point-of-care diagnostic platforms are also discussed for their potential to enhance ultra-sensitive detection, particularly in liquid biopsy settings. These emerging tools are highlighted for their potential to enable ultra-sensitive, rapid, and decentralized mutation detection, particularly in liquid biopsy settings. Key challenges, including intratumoral heterogeneity, low allele-frequency variants, FFPE-associated artifacts, and clonal evolution under therapeutic pressure, are examined within a translational framework. In addition, we examine critical barriers to clinical implementation, including standardization, cost, and global accessibility of molecular diagnostics, and outline potential solutions through scalable technologies and decentralized testing strategies. We propose that optimal BRAF testing requires a mutation subclass-informed and clinically integrated strategy combining comprehensive baseline profiling with longitudinal molecular monitoring. Future diagnostic paradigms will likely integrate multi-omics data and artificial intelligence (AI)-assisted interpretation to refine precision oncology implementation. Looking forward, we propose that optimal BRAF testing will require integration of multi-omics profiling with AI-assisted interpretation, enabling automated variant classification, real-time clinical decision support, and improved prediction of therapeutic response and resistance.

Humans

Cytonuclear conflict and reticulate evolution in the Morelloid clade (Solanum, Solanaceae): Insights from genome skimming and network Phylogenomics.

The Morelloid clade (black nightshades) is one of the most strongly supported clades within the megadiverse Solanum genus. It comprises 76 globally distributed, non-spiny herbaceous and suffrutescent species. While often erroneously considered poisonous weeds, several species are economically important as orphan crops. The clade is closely related to tomato and potato but, due to a lack of focused breeding efforts, remains a putative reservoir of genetic diversity for crop improvement. Despite this potential, we lack fundamental knowledge on the evolution of the Morelloid clade. The group includes polyploid species with unknown parental origins-likely reflecting reticulate processes such as hybridization, introgression, and associated backcrossing events. Prior analyses have been unable to disentangle these processes, leaving the mechanisms underlying reticulate evolution in the Morelloid clade poorly understood. Here, we use genome skimming to produce a well-supported maximum likelihood plastid phylogeny from complete circularized plastomes and a coalescent-based species tree from combined Angiosperms353 and conserved ortholog set nuclear markers. Our dataset, composed of previously published data and deep genome skimming from herbarium samples, spans 26 Morelloid species. To investigate phylogenetic discordance, we used a nuclear phylogenetic network, multispecies coalescent simulations, a fused rooted nuclear chloroplast tree, and quantification of nuclear gene tree concordance. We show that incongruence between nuclear and plastid trees is pervasive and cannot be explained by incomplete lineage sorting alone. Instead, our results demonstrate that events consistent with repeated chloroplast capture have shaped the reticulate evolutionary history of the clade, especially among African polyploid and Pan-American diploid lineages.

Phylogeny

Diversity and population connectivity of members of the family Eunicidae inhabiting deep-water corals in the North Atlantic.

Eunicid polychaetes are often found in association with Cold Water Corals (CWCs), even establishing symbiotic relationships, such as those described between Desmophyllum pertusum and Eunice norvegica. While genetic connectivity of CWCs across the North Atlantic has been widely studied, little is known about their associated fauna in this regard. Here, we present a study combining a focused analysis of the genetic and genomic connectivity of E. norvegica with a regional assessment of the distribution and evolutionary relationships of three CWC-associated eunicid species from the Cantabrian Sea and the North of the United Kingdom (190-1,230 m depth). An integrative approach using genetic (16S, COI and 18S), morphological and ecological data allowed the identification of the eunicids studied, with new records of Eunice cf. nicidioformis and Leodice cf. antarctica in the Cantabrian Sea, as well as previously undocumented associations with CWC species. In addition, RADseq data contributed to the delimitation of the closely related species E. norvegica and Eunice philocorallia. Moreover, the genetic connectivity of E. norvegica was studied trough a RADseq (1,067 neutral SNPs) approach. Our results indicate a single panmictic population across approximately 2,000 km, suggesting that oceanographic currents facilitate passive dispersal of E. norvegica lecithotrophic larvae, aided by coral host stepping-stones. The connectivity patterns observed for E. norvegica mirror those of D. pertusum, on which the worm is ecologically dependent. Our study highlights the importance of using integrated genetic, morphological and ecological data to characterise and delineate understudied CWC-associated species and improve our understanding of their dispersal capabilities and genetic connectivity to inform future conservation recommendations.

Animals

Whole-Exome Sequencing in a Consanguinity-Enriched South Indian Retinitis Pigmentosa Cohort: Diagnostic Yield and Molecular Spectrum.

PURPOSE: To determine the molecular diagnostic yield, variant spectrum, inheritance architecture, and influence of consanguinity on whole-exome sequencing outcomes in a South Indian retinitis pigmentosa (RP) cohort. DESIGN: Prospective, registry-based cohort study. SUBJECTS: A total of 113 affected participants were enrolled through the Aravind Registry for Inherited Diseases of the Eye, including 109 unrelated probands and 4 affected relatives from already represented families. Primary analyses were restricted to the 109 unrelated probands. METHODS: Whole-exome sequencing was performed using a clinical exome workflow. Variants were interpreted using American College of Medical Genetics and Genomics/Association for Molecular Pathology criteria and cases were categorized as solved, possibly solved, inconclusive, or unsolved using prespecified inheritance-aware rules. MAIN OUTCOME MEASURES: Molecular diagnostic yield, distribution of implicated genes and variant classes, inheritance architecture, and diagnostic yield stratified by consanguinity status. RESULTS: Among the 109 unrelated probands, mean age at testing was 39.3 ± 14.1 years and 58.7% were male. Whole-exome sequencing identified 186 distinct rare variants across 92 inherited retinal disease genes, including 26 pathogenic and 33 likely pathogenic variants. A molecular diagnosis was established in 50 of 109 probands (45.9%), including 42 solved and 8 possibly solved cases; 45 (41.3%) were inconclusive and 14 (12.8%) remained unsolved, including 4 (3.7%) in whom no candidate variant was identified. EYS, USH2A, and ADGRV1 were the most frequently implicated genes. Autosomal recessive (AR) disease predominated (44/50, 88.0%). Consanguineous AR cases were exclusively homozygous (17/17); notably, 68.0% of nonconsanguineous AR cases were also homozygous (P = 0.013). Diagnostic yield was higher in consanguineous probands (51.4% vs. 41.7%), without reaching significance. Recurrent alleles included an established South Asian founder variant (MFSD8 c.1361T>C) and candidate founder alleles in EYS (c.4321C>T) and ADGRV1 (c.14329C>T). CONCLUSIONS: Whole-exome sequencing established a molecular diagnosis in nearly half of this South Indian RP cohort and revealed a predominantly recessive, homozygosity-enriched architecture shaped by consanguinity. These findings define a region-specific variant landscape to support clinical interpretation, genetic counseling, and future trial enrollment in this underrepresented population. FINANCIAL DISCLOSURES: The authors have no proprietary or commercial interest in any materials discussed in this article.

Consanguinity

Phylogenomics and female reproductive morphology reframe the classification of the Halymeniales (Rhodophyta).

The red algal order Halymeniales (Rhodophyta) exhibits remarkable morphological and taxonomic diversity but its higher-level relationships remain poorly resolved. Here, we present a comprehensive phylogenomic analysis based on newly generated plastid (170 protein-coding genes), mitochondrial (23 genes), and complete nuclear ribosomal cistron sequences from 56 taxa, complemented with an expanded rbcL dataset encompassing 334 sequences. Our results provide a robust phylogenomic framework for the Halymeniales, offering a taxonomic backbone for future systematic studies. The analyses consistently recover six early-diverging lineages (Acrodiscus, Isabbottia, Norrissia, Pachymenia, Zymurgia, and Tsengia) and two strongly supported larger clades (Halymenia s.l. and Grateloupia s.l.). While most small and recently described genera are monophyletic, several traditional genera (e.g., Halymenia, Cryptonemia, Grateloupia) are poly- or paraphyletic, requiring considerable taxonomic revision. At the family level, the data indicate that reinstatement of the Grateloupiaceae sensu Kim et al. (2021) would entail a revised circumscription of the Halymeniaceae and the recognition of at least five small families to accommodate the early-diverging lineages. Although such a revised classification would result in monophyletic families, it is not supported by morpho-anatomical characters. Instead, we propose a more stable two-family system, recognizing a broadly circumscribed Halymeniaceae that is sister to the Tsengiaceae. Female reproductive characters, particularly the structure of carpogonial and auxiliary cell ampullae, support this two-family system and further characterize many genus-level clades, although substantial convergence across lineages exists.

Phylogeny

Cost-Effectiveness and the Economics of Genomic Testing and Molecularly Matched Therapies.

Cost-effectiveness analysis of precision oncology can help guide value-driven care. Next-generation sequencing is increasingly cost-efficient over single gene testing because diagnostic algorithms require multiple individual gene tests to determine biomarker status. Matched targeted therapy is often not cost-effective due to the high cost associated with drug treatment. However, genomic profiling can promote cost-effective care by identifying patients who are unlikely to benefit from therapy. Additional applications of genomic profiling such as universal testing for hereditary cancer syndromes and germline testing in patients with cancer may represent cost-effective approaches compared with traditional history-based diagnostic methods.

Humans

Multi-omics integrative analysis provides insight into potential molecular responses to sustained high water flow in common carp (Cyprinus carpio) cultured in recirculating aquaculture.

To investigate the potential molecular responses by which water flow intensity affects the growth of common carp (Cyprinus carpio) in a recirculating aquaculture system (RAS), a control group (CG, actual water velocity 0.3&#xa0;cm/s) and three sustained flow treatment groups were established, including a low-flow group (LF, 1 body length per second, bl/s), a medium-flow group (MF, 2 bl/s), and a high-flow group (HF, 3 bl/s). After 12&#xa0;weeks of culture in the RAS, growth performance was compared among groups under different flow intensities. The best-performing group and the control group were then selected for the determination of intestinal digestive enzyme activities, as well as transcriptomic and whole-genome bisulfite sequencing analyses of muscle tissue. The results showed that the specific growth rate and feed intake of the HF group were significantly higher than those of the other groups (P&#xa0;<&#xa0;0.05), whereas no significant difference in feed conversion ratio was observed among groups. Compared with the CG group, lipase activity was significantly higher in the HF group (P&#xa0;<&#xa0;0.05), while &#x3b1;-amylase and trypsin activities showed increasing trends without significant differences. RNA-seq identified a total of 273 differentially expressed genes, including 72 upregulated genes and 201 downregulated genes in the HF group relative to the CG group. These genes were mainly enriched in glycolysis, pyruvate metabolism, ATP metabolism, the pentose phosphate pathway, the insulin signaling pathway, the PPAR signaling pathway, and the adipocytokine signaling pathway, indicating that sustained high water flow induced a muscle transcriptional response characterized by remodeling of energy metabolism and substrate utilization. Whole-genome bisulfite sequencing analysis showed that DNA methylation in common carp muscle occurred predominantly in the CpG context. Differentially methylated regions between the HF and CG groups were mainly distributed in transcription-related regulatory regions, including promoters, CpG islands, and CpG island shores. In promoter regions, the number of hypermethylated regions in the HF group relative to the CG group was markedly higher than that of hypomethylated regions. Integrated analysis further identified two candidate genes showing both promoter differential methylation and differential expression, namely LOC109094644 and bcorl1, suggesting that adaptation to high water flow may involve IGF-related growth regulation and remodeling of upstream transcriptional programs. The qPCR results were consistent with the transcriptomic data. Taken together, within the tested range, a sustained water flow of 3 bl/s was more conducive to the growth of common carp in the RAS, which may be associated with enhanced lipid digestion and utilization, remodeling of the muscle energy metabolic network, changes in promoter methylation, and the coordinated regulation of key candidate genes. This study provides a theoretical basis for clarifying the exercise adaptation mechanism of common carp in recirculating aquaculture and for optimizing flow velocity parameters.

Animals

Integrated bioinformatics analysis reveals cross-talking hub genes and therapeutic agents between sepsis and acute myocardial infarction.

BACKGROUND: Sepsis and acute myocardial infarction (AMI) are two significant diseases that may share overlapping etiological mechanisms. This study aims to systematically identify core genes common to both conditions and to explore their potential as therapeutic targets and drug candidates through an integrative analysis of clinical data and bioinformatics. METHODS: The AMI dataset was obtained from the GEO database, and RNA sequencing data were collected from blood samples of patients with sepsis at our hospital. Common genes were identified using differential expression gene analysis (DEG) and weighted gene co-expression network analysis (WGCNA). Functional enrichment analyses, including Gene Ontology (GO) and Kyoto Encyclopedia of Genes and Genomes (KEGG) pathway analysis, were performed. A protein-protein interaction (PPI) network was constructed, and hub genes were identified using the MCC/Degree algorithm. Diagnostic value was assessed via receiver operating characteristic curve analysis. Immune infiltration patterns, single-cell sequencing data, and molecular docking simulations were employed to evaluate immune relevance and identify potential therapeutic compounds. RESULTS: A total of 417 genes were identified between sepsis and AMI, with enrichment analysis revealing significant involvement in inflammatory responses. Three hub genes-JAK2, MYD88, and TIMP1-were selected for further investigation. ROC curves confirmed their strong diagnostic performance for both diseases. Immune infiltration analysis showed that these core genes were significantly correlated with the infiltration levels of various immune cell types. Molecular docking indicated that quercetin exhibited stable binding affinity with the proteins encoded by these genes. qPCR validation further confirmed the upregulation of these three genes, supporting the anti-inflammatory effects of quercetin as a potential targeted therapy. CONCLUSION: JAK2, MYD88, and TIMP1 were identified as shared core genes in sepsis and AMI. These genes not only serve as potential diagnostic biomarkers but also offer novel targets for developing common therapeutic strategies for both conditions. Furthermore, quercetin emerges as a promising candidate for targeted treatment.

Humans

Molecular Diagnostics for WHO Priority Bacterial Pathogens: A Bibliometric Mapping of Diagnostic Platforms, Resistance Markers, and Antimicrobial Resistance Research Trends.

Antimicrobial resistance (AMR) constrains effective treatment and carries implications for infection control, surveillance, and public health. The World Health Organization (WHO) priority bacterial pathogen framework has intensified the need for diagnostic innovation by redefining research priorities around organisms combining high disease burden with complex resistance profiles. Molecular diagnostics have accordingly moved beyond culture-based workflows, integrating rapid pathogen identification, resistance-marker detection, genomic surveillance, and clinical decision support. The present study conducted a bibliometric mapping of the literature on WHO priority pathogens. Rather than addressing resistance at a general level or a single pathogen or technology, it integrates priority pathogens, molecular platforms, and resistance markers within a single framework, tracing their joint thematic and temporal evolution along an explicit pathogen-platform-marker axis. Scopus-indexed articles and reviews (2000-2025) were retrieved, yielding 1746 publications after screening adapted from the Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) guidelines. Analyses used Bibliometrix/Biblioshiny, R, and VOSviewer. The literature expanded markedly after 2018, led by China and the United States. Methicillin-resistant Staphylococcus aureus (MRSA), Mycobacterium tuberculosis, Enterococcus faecium, and the Enterobacterales-carbapenemase axis constituted the principal thematic cores, whereas conventional polymerase chain reaction (PCR)/nucleic acid amplification testing (NAAT) and whole-genome sequencing were the dominant platforms. Overall, the field has evolved from pathogen detection into an AMR-centered translational domain encompassing resistance prediction, genomic epidemiology, surveillance, and clinical decision support. Diagnostic development, stewardship, and surveillance depend on hybrid workflows coupling rapid marker-targeted assays with genome-based characterization, delivering actionable resistance within clinically meaningful timeframes, and extending coverage to underrepresented pathogens and platforms.

Humans

A genome-wide coverage-based pipeline for the identification of host-derived candidate DNA biomarkers from cell-free blood.

We have created a new data-analysis pipeline for the discovery of host-specific candidate DNA biomarkers derived from sequencing data of cell-free blood. Unlike approaches that rely on specific molecular or genetic signatures, our method leverages the coverage distribution of cell-free DNA sequences mapped to a reference genome, applying statistical analyses to identify informative short genomic regions for biomarker discovery. The pipeline is applicable to diverse diseases and can be used to analyze cell-free DNA sequences from plasma or serum to identify candidate biomarkers that are characteristic of disease states in mammals. Core functionalities were developed in Java and integrated with open-source software tools for the preprocessing of raw sequencing data, complemented by Python scripts for the machine-learning analysis and statistical validation. The pipeline is designed for HPC use and users can access the pipeline through a Galaxy workflow, which offers a user-friendly web interface for input selection prior to execution and analysis progress monitoring. Performance tests, carried out using duplicate sets of COVID-19 samples and controls, showed linear scalability of execution time with an increasing dataset size, as well as a substantial reduction in execution time through parallelized computation, whereby each HPC node is used to process the data of one chromosome. Further statistical tests confirmed the quality of the pipeline's results by showing that the set of identified candidate biomarkers remained stable across varying dataset sizes.

Biomarkers