Search PubMedSearch

subject

Machine Learning

Machine Learning: explore 9 source-linked works published from 2025 to 2026, with original documents and citations.

This collection is a preview while coverage and quality are evaluated.

Search within this collection

Coverage and selection

Includes records with this source-supplied label or an explicit phrase match in their metadata. Matches indicate a mention, not proof that a paper uses a method or tests a material. Source versions are consolidated by DOI.

Sources: pubmed. Collection updated 2026-09-15. Counts describe this index, not the complete source archives.

Machine learning-ready genomic biomarkers: ATF3 polymorphisms predict postoperative analgesic demand through AI-compatible phenotyping.

PURPOSE: To determine whether ATF3 polymorphisms can serve as genetic biomarkers for machine learning-based precision analgesia by establishing a genotype-phenotype association suitable for predictive modeling of postoperative opioid requirements. METHODS: In a prospective cohort of 167 adults undergoing abdominal surgery, ATF3 SNPs rs3122721 and rs3125293 were genotyped. A structured dataset architecture was developed to represent genetic profiles as input features for supervised learning models, enabling translational analysis of genotype‑dependent opioid consumption over 72 h. RESULTS: Patients with homozygous genotypes of the ATF3 SNPs had significantly higher opioid requirements than non‑carriers, despite reporting similar subjective pain scores. This consistent genotype‑dependent pattern provided a clinically relevant phenotype suitable for integration into predictive algorithms. CONCLUSION: ATF3 genotyping offers a promising biomarker for computationally informed precision analgesia. By linking genomic variability to clinically meaningful outcomes within a structured clinical and genomic framework, this approach supports the future development of risk-stratified clinical decision-support systems to optimize postoperative pain management.Trial registration ChiCTR1900021991, registered 30 April 2019. SUPPLEMENTARY INFORMATION: The online version contains supplementary material available at https://doi.org/10.1007/s13755-026-00480-9.

ATF3

Functional neuroimaging subtypes of obsessive-compulsive disorder: A systematic review and meta-analysis.

Obsessive-compulsive disorder (OCD) exhibits substantial clinical heterogeneity that may reflect underlying neurobiological diversity. Neuroimaging-based subtyping may advance precision psychiatry by identifying biologically distinct subgroups with differential treatment responses. This study systematically synthesized evidence from functional neuroimaging subtyping studies in OCD to identify reproducible neurobiological subtypes, characterize their clinical profiles, and establish a consensus-based classification framework. We reviewed 40 original studies employing machine learning, clustering, normative modeling, or classification approaches, encompassing approximately 8,150 patients. Consensus clustering identified three reproducible neurobiological subtypes. The Limbic-Hyperactive subtype, comprising approximately 40% of patients, exhibited amygdala and insula hyperconnectivity, elevated anxiety levels, predominant contamination and washing symptoms, and favorable response to cognitive-behavioral therapy. The Fronto-Striatal-Hypoconnected subtype, comprising approximately 35% of patients, demonstrated reduced orbitofrontal-striatal connectivity, cognitive inflexibility, predominant checking and ordering symptoms, and a favorable response to selective serotonin reuptake inhibitors. The Global-Disrupted subtype, comprising approximately 25% of patients, exhibited widespread connectivity disruption, greater symptom severity, and poor treatment response. Support vector machine classification achieved 81.5% accuracy for subtype assignment, though classification of OCD versus healthy controls showed limited generalizability in multisite settings (AUC 0.567-0.673). These findings support a neuroimaging-based framework for personalized treatment selection but require prospective validation.

Humans

Development and validation of a comprehensive prognostic model for 28-day ICU mortality in non-traumatic subarachnoid hemorrhage: an analysis based on the MIMIC-IV database.

BACKGROUND: Due to the complex pathophysiology of non-traumatic subarachnoid hemorrhage (SAH), accurate risk prediction remains a challenge. Our aim is to develop and validate a comprehensive prognostic model that integrates demographic characteristics, vital signs, laboratory parameters, and more, to provide clinical decision-making support in real-world practice. METHODS: We conducted a retrospective cohort study of 785 Non-traumatic subarachnoid hemorrhage patients. The cohort was randomly divided into a training set (n = 549) and a validation set (n = 236). Feature selection was performed using LASSO regression, followed by backward stepwise Cox regression for optimization. A nomogram was constructed based on independent predictive factors, and model performance was assessed using discrimination, calibration, and decision curve analysis. To prevent immortal-time bias, all predictors were anchored to a fixed early (first-24-hour) measurement window, treatment variables were modelled as binary indicators rather than cumulative exposures, and a five-model sensitivity analysis with baseline-severity adjustment was performed. RESULTS: The development of our model followed a systematic approach: first, 15 potential predictive factors were selected via LASSO regression, which were then refined to 12 independent predictors using backward stepwise Cox regression. The final predictive factors included: Ventilation, AHT, Nimodipine 60 mg, Age, SAPS.II, Input amount, Calcium total, Platelet count, White blood cells, Anion gap, pH, and Chloride. The integrated model demonstrated excellent predictive ability for 7-day, 14-day, and 21-day mortality in both the training set (AUC: 0.972, 0.934, 0.898) and the validation set (AUC: 0.968, 0.948, 0.911). Calibration curves and decision curve analysis confirmed the model's reliability and clinical utility across different time points. We constructed a nomogram for individualized risk prediction. Univariate Kaplan-Meier survival analysis demonstrated significant stratification of survival outcomes by each predictor, while restricted cubic spline analysis revealed non-linear relationships between continuous variables and mortality risk. Random survival forest analysis identified the top three predictive factors (Nimodipine 60 mg, Ventilation, AHT) and compared them with our full 12-variable model, confirming superior performance of the integrated model at all time points. At the 28-day primary endpoint, the model achieved a time-dependent AUC of 0.898 (training) and 0.904 (validation); after restricting predictors to the early baseline window, the leakage-controlled model retained good discrimination (validation C-index 0.803). CONCLUSIONS: Our ICU 28-day mortality prognosis model demonstrated robust performance in predicting ICU 28-day mortality in non-traumatic subarachnoid hemorrhage. The model, through the nomogram, provides individualized risk assessment, aiding clinical decision-making and patient stratification.

Humans

Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.

BACKGROUND: Predictive modelling for cancer risk, treatment-related complications, and survival is central to precision oncology. Conventional logistic regression (LR) and Cox proportional hazards (CoxPH) regression remain widely used but are limited when modelling nonlinear interactions, high-dimensional imaging features, and multimodal clinical-metabolic predictors. Artificial intelligence (AI) and machine learning (ML) methods offer expanded capability through automated feature extraction, ensemble learning, and flexible survival modelling, but the evidence on when AI/ML adds value over conventional models across cancer sites and predictive tasks remains fragmented. OBJECTIVE: To systematically evaluate the methodological performance, validation strategies, and translational limitations of AI/ML models compared with conventional statistical models in published predictive-modelling studies for breast, colorectal, or pancreatic cancer. METHODS: PubMed, Scopus, and Web of Science were searched for studies published between January 2019 and March 2025. Two reviewers independently conducted title-and-abstract screening, full-text eligibility assessment, and PROBAST risk-of-bias assessment. Sixty-five studies (n = 907,567 participants) were narratively synthesised by cancer site, predictive task, model family, comparator, validation strategy, predictor modality, and calibration or explainability reporting. RESULTS: The 65 studies comprised breast cancer (n = 35), colorectal cancer (n = 21), and pancreatic cancer (n = 9). AI/ML superiority over LR and CoxPH was task- and data-dependent. CNN- and U-Net-based models predominated in imaging and body-composition tasks, tree-based ensembles consistently outperformed LR for tabular perioperative complication prediction, and CoxPH remained competitive, and in the largest pancreatic risk study, superior to XGBoost (C-index 0.802 vs 0.723) in well-structured datasets. PROBAST analysis-domain risk was moderate in 54 of 65 studies (83%), driven by limited external validation, sparse calibration reporting (11/65), and few decision-curve analyses (7/65). CONCLUSION: AI/ML adds the most methodological value in imaging-derived feature extraction and nonlinear perioperative prediction, while conventional regression remains preferable in large, structured datasets with linear predictors. Clinical translation requires standardised body-composition definitions, external validation, calibration assessment, decision-curve analysis, and explainability, in line with TRIPOD+AI and CLAIM standards.

Humans

Diagnostic performance of machine learning models for malignant and non-malignant pleural effusion: Systematic review and meta-analysis.

BACKGROUND: Accurately distinguishing malignant pleural effusion (MPE) from non-malignant pleural effusion is clinically important, but the generalisability and methodological quality of machine-learning (ML) models remain uncertain. METHODS: We searched eight databases to 23 April 2026. Diagnostic performance was pooled using random-effects and Reitsma bivariate models, and study quality was assessed using PROBAST+AI. RESULTS: Forty-two studies were included; 17 contributed to the AUC meta-analysis and 14 to the bivariate analysis. The pooled AUC was 0.90 (95 % CI 0.85-0.94; 95 % prediction interval 0.62-0.98), with sensitivity of 0.80 (95 % CI 0.77-0.83) and specificity of 0.87 (95 % CI 0.79-0.92). Only nine studies reported external, temporal or independent validation. Externally validated studies had a lower pooled AUC than studies without external validation (0.83 vs 0.92), with lower specificity observed in the two externally validated studies contributing sensitivity and specificity data. All 42 development assessments had high overall quality concerns, and all 42 model evaluations were judged at high risk of bias. CONCLUSIONS: ML models showed good apparent accuracy for distinguishing MPE from non-MPE, but the evidence was limited by substantial heterogeneity, high risk of bias and scarce external validation. The pooled estimates reflect the average performance of different selected models rather than the expected accuracy of a single clinical test. ML models should be regarded as adjuncts to existing diagnostic pathways until they are confirmed by rigorous multicentre prospective external validation and clinical-impact studies.

Humans

Machine learning-based prediction of unplanned readmission and construction of an online calculator for elderly patients with mild ischemic stroke.

OBJECTIVE: To screen for independent risk factors for unplanned readmission in elderly patients with mild ischemic stroke, and to construct and validate an online risk prediction calculator based on an interpretable machine learning model, thereby providing a promising practical tool for accurate clinical assessment of 30&#x2011;day all&#x2011;cause unplanned readmission risk in this population. METHODS: A prospective cohort study was conducted, including 1050 patients aged&#xa0;&#x2265;&#xa0;60&#xa0;years with mild ischemic stroke admitted between August 2023 and September 2024. Participants were randomly divided into a training set (840 cases) and a test set (210 cases) at a ratio of 8:2. Risk factors were screened by univariate analysis and multivariable Logistic regression. Four machine learning models, namely LightGBM, XGBoost, Random Forest, and K&#x2011;Nearest Neighbors (KNN), were developed and their performance was evaluated using AUC, accuracy, sensitivity, and specificity as metrics. The SHAP framework was used for interpretability analysis, and an online calculator was subsequently developed based on the optimal model. RESULTS: Univariate analysis showed significant differences (P&#xa0;<&#xa0;0.05) in 13 factors including age, smoking, AIP, TyG index, HALP score, etc. Multivariable Logistic regression identified age (OR&#xa0;=&#xa0;9.752), smoking (OR&#xa0;=&#xa0;5.171), AIP (OR&#xa0;=&#xa0;6.691), TyG index (OR&#xa0;=&#xa0;4.393), HALP score (OR&#xa0;=&#xa0;2.831), and&#xa0;&#x2265;&#xa0;2 comorbidities (OR&#xa0;=&#xa0;3.664) as independent risk factors. All four machine learning models demonstrated good predictive performance. Based on a comprehensive evaluation of multiple metrics and computational efficiency, the LightGBM model exhibited the best predictive performance (AUC&#xa0;=&#xa0;0.884, accuracy&#xa0;=&#xa0;0.829, sensitivity&#xa0;=&#xa0;0.812, specificity&#xa0;=&#xa0;0.875). SHAP analysis showed that age, AIP, TyG index, smoking, and HALP score were key predictors. An online calculator developed based on this model enables individualized risk predictions. CONCLUSION: Key risk factors associated with 30&#x2011;day unplanned readmission in elderly patients with mild ischemic stroke were identified. The LightGBM model demonstrated high predictive accuracy, and together with the interpretability analysis and online calculator, offers a practical tool to support clinical risk assessment. However, this tool requires future external validation.

Humans

Integrative machine learning and transcriptomic analysis reveals molecular mechanisms underlying low survival rate in larval Chinese Bahaba (Bahaba taipingensis).

Chinese Bahaba (Bahaba taipingensis) is a Class I protected marine fish endemic to China. Low larvae survival during artificial breeding severely hinder population recovery. To investigate the molecular mechanism of high mortality in larval fish, this study performed RNA-seq on liver from naturally deceased (ND) and mass-dead (MD) individuals, combined with least absolute shrinkage and selection operator (LASSO) regression and random forest (RF) algorithms to screen for core signature genes. A total of 873 differentially expressed genes (DEGs) were identified, including 112 upregulated and 761 downregulated genes. GO and KEGG enrichment analyses revealed significant enrichment in amino acid metabolism disorders, one&#x2011;carbon folate pool impairment, PPAR signaling abnormalities, ECM-receptor interaction, focal adhesion pathway, indicating widespread metabolic suppression accompanied by extracellular matrix remodeling and signaling disturbances in the livers of MD fish. MAD pre-filtering combined with dual machine learning algorithms yielded 18 robust core signature genes, among which SLC38A4, MMP1, FADD, FKBP5, and APOB were consistently identified as high-frequency core genes by both algorithms. SLC38A4 exhibited the highest importance score in the RF model and was significantly downregulated, making it the primary molecule distinguishing ND from MD phenotypes. ROC curve analysis showed that both models achieved an AUC of 1.000 (95% CI lower bound: 0.610), confirming the precise discriminatory ability of the core genes. GSEA further demonstrated significant enrichment of this core gene set in ND samples. This study provides the first systematic elucidation of the molecular mechanisms underlying liver dysfunction in low survival rate B. taipingensis, characterized by amino acid transport impairment, metabolic reprogramming, and structural remodeling, offering theoretical foundations for health assessment, early mortality risk warning, and artificial breeding conservation of this species.

Animals

Nature-based meaning-focused photography intervention enhances subjective well-being: A three-arm randomized controlled study.

Gaining meaning from nature contact can promote subjective well-being. However, few studies have validated the effectiveness of nature-based meaning interventions in enhancing subjective well-being. This study consisted of a 7-day online intervention to examine the effects of nature-based meaning-focused photography on well-being by comparing a photo-only group, a photo&#x2009;+&#x2009;writing group, and a waiting list control group and how meaning in life mediates the relationship between nature contact and well-being. A pre-registered three-arm randomized controlled trial (groups: photo&#x2009;+&#x2009;writing group vs. photo-only group vs. control group)&#xa0;*&#xa0;(time: pre-test vs. post-test vs. 1-month follow-up) was conducted with 219 college students. In the photo&#x2009;+&#x2009;writing group, participants captured nature scenes and wrote 100-word reflections. The photo-only group only took nature photos. The primary outcomes were meaning in life and well-being, and the secondary outcome was life satisfaction. A conservative Bayesian causal forest analysis based on machine learning was used to detect both treatment and heterogeneous intervention effects. Compared with the control group, the photo&#x2009;+&#x2009;writing group showed positive effects on meaning in life, subjective well-being, and life satisfaction, with average treatment effects of 0.36, 0.27, and 0.66 standard deviations (SD), respectively. The photo-only group also showed generally positive effects on these outcomes, with average treatment effects of 0.27, 0.24, and 0.54 SD, respectively. However, these effects were not sustained after 1&#x2009;month. The intervention was especially beneficial for participants from lower subjective socioeconomic status, with limited prior nature exposure, or lower baseline psychological well-being. Importantly, enhanced meaning in life helped explain how the intervention improved well-being and life satisfaction. This study also demonstrated that combining nature-based photography and reflective writing can improve well-being.

Humans

Application of causal discovery of factors driving dissolved oxygen in estuarine environments.

Dissolved oxygen (DO) concentrations in estuarine bottom waters are a manifestation of multiple, interacting physical and biogeochemical processes, yet identifying their independent contributions remains challenging. Here, we analyze monthly water quality monitoring data from eight stations across Long Island Sound from 1994 to 2022 using a causal discovery framework (PCMCI+) and transformation of forcing variables. Our goal is to identify and isolate variables that causally influence bottom DO and improve predictive models by minimizing overfitting and multicollinearity. PCMCI+ reveals surface-layer temperature as the most important and consistent negative driver of bottom DO, followed by stratification. Wind events exhibit only brief relief by advection and mixing, while river discharge shows no direct causal link to DO, making it less influential than previously thought. Biogeochemical variables, including chlorophyll-a (Chl-a), nitrate and nitrite, and particulate carbon, influence DO through both contemporaneous and time-lagged pathways, often with signs that shift depending on the process. The derived models were evaluated by comparing skill scores, mean squared error, and Akaike Information Criterion. Both model types perform well, with coefficient of determination values exceeding 0.90 at multiple stations using only 3-5 predictors. Our analysis reveals that the best causal predictors are surface-layer temperature, stratification, Chl-a, and particle carbon. This approach provides a scalable framework for improving prediction models and understanding the mechanistic links that control the seasonal variability of DO in estuarine systems.

Estuaries
Compare source metadata on this page
WorkPublishedSource identifierSource
Machine learning-ready genomic biomarkers: ATF3 polymorphisms predict postoperative analgesic demand through AI-compatible phenotyping.2026-09-02PMID 42694296pubmed
Functional neuroimaging subtypes of obsessive-compulsive disorder: A systematic review and meta-analysis.2026-08-22PMID 42648054pubmed
Development and validation of a comprehensive prognostic model for 28-day ICU mortality in non-traumatic subarachnoid hemorrhage: an analysis based on the MIMIC-IV database.2026-08-19PMID 42617498pubmed
Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.2026-08-17PMID 42641254pubmed
Diagnostic performance of machine learning models for malignant and non-malignant pleural effusion: Systematic review and meta-analysis.2026-08-16PMID 42617201pubmed
Machine learning-based prediction of unplanned readmission and construction of an online calculator for elderly patients with mild ischemic stroke.2026-08-05PMID 42556003pubmed
Integrative machine learning and transcriptomic analysis reveals molecular mechanisms underlying low survival rate in larval Chinese Bahaba (Bahaba taipingensis).2026-07-18PMID 42480247pubmed
Nature-based meaning-focused photography intervention enhances subjective well-being: A three-arm randomized controlled study.2026PMID 42725703pubmed
Application of causal discovery of factors driving dissolved oxygen in estuarine environments.2025-11-14PMID 42680405pubmed

These are bibliographic comparisons, not experimental rankings. Follow the original document for methods and conditions.