Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Machine learning.”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,621 records · Page 90Linked to original sources

Reversal of cancer gene expression identifies repurposed drugs for diffuse intrinsic pontine glioma.

Diffuse intrinsic pontine glioma (DIPG) is an aggressive incurable brainstem tumor that targets young children. Complete resection is not possible, and chemotherapy and radiotherapy are currently only palliative. This study aimed to identify potential therapeutic agents using a computational pipeline to perform an in silico screen for novel drugs. We then tested the identified drugs against a panel of patient-derived DIPG cell lines. Using a systematic computational approach with publicly available databases of gene signature in DIPG patients and cancer cell lines treated with a library of clinically available drugs, we identified drug hits with the ability to reverse a DIPG gene signature to one that matches normal tissue background. The biological and molecular effects of drug treatment was analyzed by cell viability assay and RNA sequence. In vivo DIPG mouse model survival studies were also conducted. As a result, two of three identified drugs showed potency against the DIPG cell lines Triptolide and mycophenolate mofetil (MMF) demonstrated significant inhibition of cell viability in DIPG cell lines. Guanosine rescued reduced cell viability induced by MMF. In vivo, MMF treatment significantly inhibited tumor growth in subcutaneous xenograft mice models. In conclusion, we identified clinically available drugs with the ability to reverse DIPG gene signatures and anti-DIPG activity in vitro and in vivo. This novel approach can repurpose drugs and significantly decrease the cost and time normally required in drug discovery.

Humans↗

Benchmarking with synthetic communities provides a baseline for virus-host inferences from Hi-C proximity linking.

Microbiomes influence diverse ecosystems, and viruses increasingly appear to impose key constraints. While viromics has expanded genomic catalogs, host identification for these viruses remains challenging due to the limitations in scaling cultivation-based approaches and the uncertain reliability and relative low resolution of in silico predictions - particularly for understudied viral taxa. Towards this, Hi-C proximity ligation uses sequenced, cross-linked virus and host genomic fragments to infer virus-host linkages and has now been applied in at least 10 studies. However, its accuracy remains unknown. Here we assess Hi-C performance in recovering virus-host interactions using synthetic communities (SynComs) composed of four marine bacterial strains and nine phages with known interactions and then apply optimized bioinformatic protocols to natural soil samples. In SynComs, standard Hi-C sample preparations and analyses showed poor normalized contact score performance (26% specificity, 100% sensitivity, incorrect matches up to class level) that could be dramatically improved by Z-score filtering (Z ≥ 0.5, 99% specificity), though at reduced sensitivity (62% down from 100%). Detection limits were established as reproducibility was poor below minimal phage abundances of 105 PFU/mL. Applying optimized bioinformatic protocols to natural soil samples, we compared virus-host linkages inferred from proximity-ligated Hi-C sequencing with predictions generated by in silico homology-based and machine learning-based bioinformatic approaches. Prior to Z-score thresholding, agreement was relatively high at the phylum to family levels (72%), but not at the genus (43%) or species (15%) levels. Z-score thresholding reduced sensitivity (only 34% of predictions were retained), with only modest improvements in congruence with bioinformatic methods (48% or 18% at genus or species levels, respectively). Regardless, this led to 79 genus-level-congruent virus-host linkages and 293 new ones revealed by Hi-C alone, i.e., providing many new virus-host interactions to explore in already well-studied climate-critical soils. Overall, these findings provide empirical benchmarks and methodological guidelines to improve the accuracy and reliability of Hi-C for virus-host linkage studies in complex microbial communities.

Benchmarking↗

Thinking the impossible: how to solve the protein folding problem with and without homologous structures and more.

Structure prediction of proteins is a difficult task as well as prediction of protein-protein interaction. When no homologous sequence with known structure is available for the target protein, search of distantly related proteins to the target may be done automatically (fold recognition/threading). However, there are difficult proteins for which still modeling on the basis of a putative scaffold is nearly impossible. In the following, we describe that for some specific examples, human expertise was able to derive alignments to proteins of similar function with the aid of machine learning-based methods specifically suited for predicting structural features. The manually curate search of putative templates was successful in generating low-resolution three-dimensional (3D) models in at least two cases: the human tissue transglutaminase and the alcohol dehydrogenase from Sulfolobus solfataricus. This is based on the structural comparison of the model with the 3D protein structure that became available after prediction. For protein-protein interaction, a knowledge-based method can give predictions of putative interaction patches on the protein surface; this feature may help in adding additional weight to specific nodes in nets of interacting proteins.

Alcohol Dehydrogenase↗

Emerging genes implicated in human congenital heart disease: a 2023-2025 scoping review.

BACKGROUND: Congenital heart disease (CHD) is the most common major congenital anomaly and a leading cause of infant morbidity and mortality. The rapid expansion of genomic technologies has accelerated the discovery of rare genetic variants implicated in CHD pathogenesis. However, most individuals with CHD still lack an identifiable molecular etiology. The purpose of this scoping review is to systematically characterize genes reported in the recent literature as candidate CHD-associated genes and contextualize these findings within the stages of cardiac morphogenesis. METHODS: PubMed was searched using predefined terms related to CHD and genetic variants, supplemented by a prospectively maintained internal database. We included human studies published between January 2023 and December 2025 that identified pathogenic, likely pathogenic, or uncertain monogenic variants in at least one patient with CHD. Animal-only studies, chromosomal abnormalities, copy number variants, multigenic associations, transcriptomic/proteomic analyses, reviews, and maternal-only genetic studies were excluded. Gene-disease validity classifications were assigned using the Clinical Genome Resource (ClinGen) CHD Gene Curation Expert Panel framework. RESULTS: Of 2,834 screened articles, 391 studies met inclusion criteria, identifying 912 unique genes reported as candidate CHD-associated genes. Frequently reported genes included PTPN11, NOTCH1, GATA4, JAG1, MYH6, GATA6, and LZTR1. Identified genes spanned all major stages of cardiogenesis, including developmental priming, cardiac progenitor specification, left-right axis formation, neural crest migration, outflow tract development, septation, and postnatal structural remodeling. Studies increasingly implicated ciliary dysfunction, transcriptional regulation, ribosomal biology, and multigenic inheritance in CHD pathogenesis. Emerging methodologies included stem cell-derived cardiac models, machine learning-based gene prioritization, and epigenetic analyses. CONCLUSIONS: Recent literature substantially expands the catalog of candidate genes that may be associated with CHD and highlights the biologic complexity underlying cardiac morphogenesis. Integration of genomic, developmental, and functional approaches will be essential to improve mechanistic understanding, refine genetic counseling, and support future precision medicine strategies for CHD.

Cardiac development↗

Adeno-Associated Virus Engineering and Load Strategy for Tropism Modification, Immune Evasion and Enhanced Transgene Expression.

Gene therapy aims to add, replace or turn off genes to help treat disease. To date, the US Food and Drug Administration (FDA) has approved 14 gene therapy products. With the increasing interest in gene therapy, feasible gene delivery vectors are necessary for inserting new genes into cells. There are different kinds of gene delivery vectors including viral vectors like lentivirus, adenovirus, retrovirus, adeno-associated virus et al, and non-viral vectors like naked DNA, lipid vectors, polymer nanoparticles, exosomes et al, with viruses being the most commonly used. Among them, the most concerned vector is adeno-associated virus (AAV) because of its safety, natural ability to efficiently deliver gene into cells and sustained transgene expression in multiple tissues. In addition, the AAV genome can be engineered to generate recombinant AAV (rAAV) containing transgene sequences of interest and has been proven to be a safe gene vector. Recently, rAAV vectors have been approved for the treatment of various rare diseases. Despite these approvals, some major limitations of rAAV remain, namely nonspecific tissue targeting and host immune response. Additional problems include neutralizing antibodies that block transgene delivery, a finite transgene packaging capacity, high viral titer used for per dose and high cost. To deal with these challenges, several techniques have been developed. Based on differences in engineering methods, this review proposes three strategies: gene engineering-based capsid modification (capsid modification), capsid surface tethering through chemical conjugation (surface tethering), and other formulations loaded with AAV (virus load). In addition, the major advantages and limitations encountered in rAAV engineering strategies are summarized.

Dependovirus↗

Transcriptome Analysis and Experimental Validation of Palmitoylation- Related Biomarkers in Atherosclerosis.

INTRODUCTION: Protein palmitoylation contributes to membrane localisation, signal transduction, and cell-fate regulation. It is closely associated with lipid metabolic dysfunction, immune inflammation, and vascular remodelling in atherosclerosis (AS). However, key palmitoylation-related transcriptomic markers and their potential causal associations with AS remain incompletely defined. METHODS: The Gene Expression Omnibus (GEO) dataset GSE100927 was used as the training cohort, and GSE43292 was used as an external validation cohort. Differentially expressed genes were identified using limma and intersected with palmitoylation-related genes to obtain palmitoylation-related differentially expressed genes (PRDEGs). Gene Ontology (GO) and Kyoto Encyclopedia of Genes and Genomes (KEGG) enrichment analyses were then performed using clusterProfiler. Two-sample Mendelian randomisation was used to evaluate potential causal relationships between characteristic genes and AS. Feature selection was conducted using random forest and support vector machine recursive feature elimination (SVM-RFE), and the overlapping genes selected by both methods were retained. Receiver operating characteristic (ROC) curves were used to assess diagnostic performance. A five-gene nomogram was constructed, and its clinical utility was evaluated using calibration curves and decision curve analysis (DCA). Gene set variation analysis (GSVA) was applied to compare pathway activity between high- and low-expression groups for each core gene. Single-cell analysis using Seurat and expression-based cell-cell communication analysis using CellChat were conducted with GSE159677, and upstream transcription factors were predicted using NetworkAnalyst. For in vivo validation, an AS model was established in ApoE⁸/⁸ mice fed a high-fat diet, and aortic gene and protein expression were assessed by RT-qPCR and western blotting. RESULTS: In GSE100927, 51 PRDEGs were identified. GO and KEGG enrichment analyses highlighted pathways associated with regulation of monoatomic ion transport, sarcomere and myofibril organisation, and immune inflammation. Mendelian randomisation suggested a potential protective causal association between SLC7A7 and AS. By integrating MR with random forest and SVM-RFE feature selection, we prioritised five core genes: PLCB2, GMIP, NEXN, PLN, and SLC7A7. These genes showed good diagnostic performance in GSE43292. The resulting nomogram was well calibrated and demonstrated stable net benefit in decision curve and clinical impact curve analyses. Single-gene GSVA identified consistently activated pathways across multiple genes, including innate and adaptive immune recognition, calcium signalling and myocardial contraction/cardiomyopathy, extracellular matrix-receptor interaction, cell junction pathways, autophagy-lysosome pathways, and several metabolic programmes. At the single-cell level, PLCB2 and GMIP were predominantly expressed in T cells and macrophages, NEXN and PLN were enriched in vascular smooth muscle cells, and SLC7A7 was mainly expressed in macrophages. CellChat analysis indicated increased signals for immune-related ligand-receptor interactions. In ApoE⁸/⁸ mice fed a high-fat diet, PLCB2, GMIP, and SLC7A7 were upregulated, whereas NEXN and PLN were downregulated; protein-level changes were concordant with the transcriptomic trends. DISCUSSION: These findings indicate that palmitoylation-related dysregulation in AS converges on immune inflammation, calcium signalling/contractile programmes, ECM remodelling, and autophagy-linked metabolism. The five-gene panel is supported by external validation, single-cell localisation to immune and vascular compartments, and concordant results in ApoE⁸/⁸ mice. CONCLUSION: This study identified and validated five palmitoylation-related genes associated with AS. SLC7A7 showed a potential protective causal signal in MR analysis. The enriched pathway patterns linked these genes to immune inflammation, calcium signalling-contraction coupling, ECM remodelling, cell adhesion, and autophagy- associated metabolic reprogramming. The five-gene nomogram showed potential utility for diagnostic classification and decision support, nominating candidate biomarkers and pathway targets for AS molecular subtyping, diagnosis, and mechanistic investigation.

Atherosclerosis (AS)↗

How AI Is Speeding Up the Diagnostic Odyssey for Rare Diseases.

The road to diagnosis can be long and sometimes unending for rare diseases, requiring training and resources that many clinics do not have. In this News and Perspectives article, JMIR Correspondent Simon Spichak reports on how AI initiatives at a children's hospital in the United States and one in Canada are helping bridge that gap and could fundamentally reshape the diagnostic experience for children and families living with rare diseases.

Rare Diseases↗

Human-AI Interaction With AI-Assisted Tumor Overlays in Pediatric Whole-Body Magnetic Resonance Imaging: Exploratory Reader Study.

BACKGROUND: AI tools have the potential to enhance personalized clinical care, particularly in radiology. However, their integration into clinical workflows remains complex, especially in pediatric oncology, where early cancer detection is critical. Children with Li-Fraumeni syndrome (LFS), a rare cancer predisposition disorder, undergo regular surveillance whole-body magnetic resonance imaging (wbMRI), which presents an opportunity for AI-assisted tumor detection. OBJECTIVE: We evaluated the feasibility of an AI-assisted overlay for highlighting tumor-like regions in pediatric surveillance wbMRI and explored how access to the overlay influenced radiologist workflow, candidate-lesion marking behavior, follow-up recommendations, and perceived workload. METHODS: We developed a patch-based AI segmentation model trained on augmented 2D slices from 675 surveillance wbMRI volumes of pediatric patients with LFS. The model was designed to highlight regions with high tumor probability. A reader study was conducted with 2 radiologists who independently reviewed wbMRI cases both with and without AI assistance. We measured evaluation time, number and location of reader-marked candidate lesions, type of follow-up recommendation, and subjective feedback using structured questionnaires. RESULTS: AI assistance altered interpretation workflows for both radiologists, with mixed effects. On average, the time required to evaluate each case increased when using the AI tool for both radiologists. However, one radiologist had an increase in the number of candidate lesion locations selected with the tool, and one had a decrease in the number of candidate lesion locations selected with the tool. Subjective feedback indicated that one of the radiologists reported lower mental demand with the AI tool, while both radiologists reported lower stress with the AI tool. Interrater variability was evident, underscoring the need for personalized calibration of AI tools. CONCLUSIONS: AI-assisted wbMRI interpretation can improve tumor detection in pediatric cancer surveillance by reducing false negatives. However, its influence on workflow efficiency and interradiologist variability highlights the importance of careful implementation. Successful integration requires addressing challenges such as improving the predictive precision of AI models, offering intuitive end-user designs and instructions, and building trust in AI outputs. AI outputs can influence workflow and behavior in reader-specific ways. Clinical translation will require larger, randomized, multireader studies and model refinement to reduce false positives and quantify lesion-level reader performance. This can help ensure better patient outcomes in addition to reduced clinician burnout.

Humans↗

Adverse Experiences in Brief Meditation Practices: Randomized Controlled Trial.

BACKGROUND: Meditation has become increasingly popular in recent decades. However, relatively little remains known about the prevalence of and risk factors for adverse experiences related to a single meditation practice. OBJECTIVE: The objective of our study was to examine adverse experiences associated with 3 brief, digitally delivered meditation practices (mindfulness, self-compassion, and gratitude) relative to using the internet as usual, as well as to investigate whether preintervention characteristics could predict such outcomes. METHODS: In a secondary analysis of a randomized controlled trial using samples that were representative of the US and UK adult populations with regard to ethnicity, sex, and age, we examined adverse experiences associated with 3 brief (ie, 5 or 10 minutes) meditation practices (ie, mindfulness, self-compassion, and gratitude) relative to using the internet as usual. We also investigated the potential of using preintervention characteristics to predict such outcomes. RESULTS: A total of 5049 participants completed all preintervention measures and were randomly assigned to meditation or control conditions. Across the sample, 4.1% (204/4925) of participants reported having a distressing experience during the intervention, and 7.1% (348/4908) of participants experienced an increase in negative affect from before to after the intervention. The results showed that participants who were randomized to a brief meditation intervention were no more likely to report a distressing experience than those who were randomized to use the internet as usual (odds ratio [OR] 1.05, 95% CI 0.76-1.47; P=.76). The results also showed that participants who were randomized to a brief meditation intervention were less likely to report clinically relevant increases in negative affect relative to using the internet as usual (OR 0.63, 95% CI 0.50-0.80; P<.001). Notably, participants in the 10-minute condition had a significantly higher likelihood of reporting a distressing experience than those in the 5-minute condition (OR 1.42, 95% CI 1.07-1.89; P=.02). Preintervention characteristics showed acceptable discrimination ability to predict a distressing experience (area under the curve=0.73) and slightly lower ability to predict increased negative affect (area under the curve=0.67). CONCLUSIONS: Taken together, we found that the brief, digitally delivered meditation practices tested in this study carry risks of adverse experiences that are comparable to or lower than those of typical activities on the internet; 10-minute condition was more likely to result in distressing experiences than 5-minute condition; and adverse responses to a brief meditation practice can, at least to a certain degree, be predicted using preintervention characteristics. TRIAL REGISTRATION: Open Science Framework 94HKS; https://osf.io/94hks/overview.

Humans↗

Multi-omics identification and functional validation of signal regulatory protein gamma as a prognostic biomarker and immune regulator in head and neck squamous cell carcinoma.

BACKGROUND: Head and neck squamous cell carcinoma (HNSCC) comprises biologically diverse tumors, and durable responses to immune-checkpoint blockade are achieved by only a subset of patients. There remains a need for markers that connect clinical outcome with malignant-cell phenotypes and tissue-level immune organization. METHODS: We integrated The Cancer Genome Atlas HNSCC cohort (TCGA-HNSC), five Gene Expression Omnibus (GEO) validation cohorts, single-cell RNA sequencing, Visium spatial transcriptomics, cellular indexing of transcriptomes and epitopes by sequencing (CITE-seq)-informed protein-potential inference, pharmacogenomic screening, genetic-risk analysis and experimental validation. A reconstructed 296-pipeline survival modelling framework was used to prioritize prognostic hub genes across validation-cohort-specific analyses. RESULTS: SIRPG was repeatedly ranked among the top ten selected genes in all five validation cohorts. At single-cell resolution, SIRPG-high tumor cells showed stronger malignant-cell features, immune-inhibitory and metabolic programs, Scissor-positive risk association, CLCA2/P53-related perturbation signals and inferred SIRPG-CD47/signal regulatory protein (SIRP) communication. Spatial analyses placed this axis within an immune-checkpoint-coupled niche, supported by Maxspin/multiview intercellular spatial modelling (MISTy) spatial coupling, communication analysis by optimal transport (COMMOT)-inferred CD47-SIRPG communication and scProTrans-inferred CD47/SIRPG protein-potential overlap. Functionally, SIRPG knockdown reduced HNSCC cell viability and increased apoptosis, whereas re-expression of short hairpin RNA (shRNA)-resistant SIRPG restored the CLCA2-BAX/BCL2 protein response. CONCLUSION: Together, these findings identify SIRPG as an immune-related prognostic hub and context-dependent tumor-cell regulator associated with apoptosis, immune communication and spatial microenvironmental organization in HNSCC.

Humans↗

Evaluating transmembrane topology prediction methods for the effect of signal peptide in topology prediction.

Reported performance of existing transmembrane (TM) topology prediction methods were often based on evaluations which neglected the risk of signal peptides (SP) being predicted as putative TM as well. Here, we evaluated 12 selected TM topology prediction methods (TMpred, TopPred II, DAS, TMAP, MEMSAT 2, SOSUI, PRED-TMR2, TMHMM 2.0, HMMTOP 2.0, SPLIT 3.5, TM Finder, and MPEx) for the effect of SP in prediction performance considering three SP treatments, namely: "remain" (untreated), "removed first", and "removed later". The results showed that the presence of SP significantly affected the prediction performance of the 12 selected TM topology prediction methods for all three predicted attributes (the number of transmembrane segments (TMSs), the number of TMSs plus position, and the N-tail location) and for the predicted topology (combined predictions of three attributes) by causing a reduction in prediction accuracy. In particular, lower prediction accuracies were obtained if SP is left untreated (remain) while significant increases were observed if SP is removed either first or later. However, between "removed first" and "removed later" SP treatments, the difference was statistically insignificant. In addition, we found that machine learning-based prediction methods were less affected by the presence of SP than hydropathy-based methods, but still the potential risk of degrading the prediction performance is there however to a lesser degree. Thus, when performing genome-wide analysis, the SP issue should be addressed during TM topology prediction.

Algorithms↗

Contextual quick-learning and generalization by humans and machines.

In a previous study (1994 Network: Comput. Neural Syst. 5 203-27) we compared human quick-learning and generalization (quick modelling) with that of neural nets (feedforward architectures), symbolic algorithms (decision tree procedures), and pattern classifiers (truth-set descriptors). Those studies raised the question of the role of context in the nature and rapidity of human learning. Here we address that issue in the setting of the same basic experiment (Quinlan classification problem) used for the previous studies. A major implication of our findings is that humans overwhelmingly seek, create, or imagine context in order to provide meaning when presented with abstract or apparently incomplete or contradictory or otherwise untenable situations.

Adolescent↗