Search PubMedSearch

SEARCH · Search PubMed

Results for “mortality prediction”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

517 records · Page 5Linked to original sources

Assessing comorbidities and predicting risk: A primer for APRNs.

Today's clinical environments are rife with tools designed to comprehensively account for medical complexity and comorbidities while predicting risk for a host of adverse health-related outcomes. Therefore, it is imperative that advanced practice registered nurses (APRNs) understand the structure and function of these tools, their similarities and differences, their limitations, and strategies for appropriate incorporation into practice. This article offers a practical overview for APRNs, emphasizing clinical implications and guidance for aligning assessment tools with the clinical population of interest to improve care delivery, quality, and patient outcomes.

Humans

Construction of precision clinical-proteomics risk model based on machine learning for predicting heart failure in type II diabetes mellitus.

BACKGROUND AND AIMS: Heart failure (HF) is a severe complication in type 2 diabetes mellitus (T2DM), but current risk stratification scores have limited predictive accuracy. We aimed to develop novel prediction tools integrating clinical variables with proteomics to improve risk stratification of hospitalization for HF in T2DM. METHODS AND RESULTS: In this study, we included 2111 UK Biobank participants with T2DM but no prior HF, and profiled 2920 proteins to predict 10-year incident HF hospitalization. Participants were randomly divided into training (70%), tuning (10%), and validation (20%) sets.Three prediction models were developed: a Clinical model based on demographic characteristics, comorbidities, medication use, and laboratory indices; a Protein model based on 40 proteins selected by the Light Gradient Boosting Machine (LGBM); and the Clinical OMics and Protein ASSessment for Heart Failure (COMPASS-HF) model, which integrated both clinical variables and the LGBM-selected proteins. Models were evaluated for area under the curve (AUC), sensitivity, and specificity. During follow-up, 168 participants (7.96%) developed incident HF. The COMPASS-HF model showed better discrimination than the Clinical model, with an AUC of 0.897 (95% CI: 0.850-0.945) versus 0.790 (95% CI: 0.723-0.856). It also demonstrated higher sensitivity (0.882; 95% CI: 0.725-0.967) and consistent performance in subgroups. COMPASS-HF effectively stratified risk of hospitalization for HF, with cumulative incidence rates of 31.9% in the high-risk group and 1.2% in the low-risk group. CONCLUSIONS: By combining clinical and proteomic variables, we developed a high-performance HF prediction model for T2DM, enabling precise risk stratification and informing early intervention strategies.

Humans

Predictive evolutionary genomics: principles, validation, and practice.

Climate change and habitat loss are driving rapid evolutionary responses in populations world-wide, which creates an urgent need for evolutionary forecasting in conservation and agriculture. Such forecasting can be categorized into three time scales: trait-based models that use multivariate quantitative genetic equations to project correlated phenotypic responses up to c. 20 generations, allele-based analyses that model allele frequency dynamics up to 100 generations, and composite adaptation scores that aggregate many small effects to yield predictions across longer horizons. However, these approaches have remained largely disconnected. Here, we present a Bayesian framework that integrates these three complementary approaches for evolutionary prediction. Our framework combines genomic, phenotypic, and environmental data to yield probabilistic predictions with explicit uncertainty. We show how predictive evolutionary forecasts can be validated with experimental evolution, field experimentation, historical specimens, and reciprocal transplants. These validated forecasts can help advance conservation and agricultural programmes by helping predict which populations are at risk of future extinction, optimizing breeding programmes for future climates, and planning ecosystem management under environmental change. By supporting a shift towards more predictive approaches in evolutionary biology, this framework may help improve our ability to manage biodiversity and food security in a changing world.

Genomics

Predictive Validity of Violence Screening Tools in Emergency and Psychiatric Services: A Systematic Review.

Violence against healthcare staff, including a threat or an act of violence toward people during their work, poses a physical and psychological risk to workers internationally. Screening is an important strategy in preventing violence against healthcare professionals. The aim of this systematic review was to synthesize evidence on the predictive validity of risk assessment tools used to screen for violence and aggression risk toward healthcare workers in emergency and psychiatric departments (PD). Primary studies that examined the predictive validity of risk assessment tools for workplace violence were identified via a systematic search of Medline, PsycINFO, Embase, and the Cochrane databases. There were 62 eligible studies, ten of which had a lower risk of bias (RoB). Those studies with high RoB were primarily due to a failure to present calibration measures as part of the analysis. All included studies adopted a longitudinal design and were conducted in PDs. The ten highest-quality studies reported on eight different instruments, four of which showed acceptable to outstanding predictive performance. The Dynamic Appraisal of Situational Aggression and the Brøset Violence Checklist showed the best predictive performance; they were also validated in emergency departments and are best suited for short-term risk prediction. We recommend that the selection of a risk assessment tool should consider the following: (a) the target population, (b) the violence operationalization, and (c) the purpose of the monitoring. We note that the use of a screening tool should be a part of a multicomponent strategy to ensure staff safety.

Humans

Future promise, current clinical ambiguity: a systematic review of machine learning algorithm outputs predicting risk of cardiovascular disease.

OBJECTIVE: To examine whether the outputs of machine learning algorithms designed to predict risk of cardiovascular disease (CVD) address known deficiencies of the Framingham Risk Score (FRS) and improve risk estimates. METHODS: For this critical review, Medline, Embase and IEEE were searched from inception to 1 January 2025. Included were studies describing machine learning algorithms designed to specifically compare output of cardiovascular risk assessment with the FRS. Commentaries, letters, unpublished work or non-peer-reviewed papers were excluded.Following Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) guidelines, two reviewers screened titles and abstracts independently, then populated a purpose-built data extraction form. A subsequent qualitative thematic analysis focused on algorithms' strengths, added value, potential harms, unintended consequences and equity implications.The main outcome assessed was whether, among healthy adults, the algorithm improved CVD risk prediction relative to the FRS. RESULTS: Of 707 studies retrieved, 29 met inclusion criteria. 23 reported improved predictive ability relative to the FRS. Most datasets and/or medical records used included sociodemographic predictors of CVD not included among FRS inputs. Some added costly diagnostic tests like CT angiography to FRS screening indicators. When they were defined, inputs and outcomes such as hypertension or myocardial infarction did not always adhere to FRS values. Statistical significance was generally taken as a proxy for clinical significance. Some algorithms overestimated the number at risk compared with the FRS without discussing whether that larger proportion might be at risk of overdiagnosis rather than CVD, while a few decreased the proportion found to be at risk. CONCLUSIONS: Use of artificial intelligence to improve accuracy of risk assessment for CVD demonstrates the technological capacity to merge known sociodemographic predictors with biologic variables and examine non-linear interactions among these. Still needed to achieve patient benefit is clinical insight, adherence to screening principles and cost-benefit assessment of inputs selected.

Humans

Surgical management of esophageal atresia with tracheoesophageal fistula in extremely low birth weight neonates: A systematic review.

BACKGROUND: Surgical management of esophageal atresia/tracheoesophageal fistula (EA/TEF) in extremely low birth weight (ELBW) neonates remains challenging and controversial. This study systematically reviews surgical strategies and outcomes in this population. METHODS: Following PRISMA guidelines, Cochrane, Embase, MEDLINE, Scopus, and Web of Science (2004-2024) were searched in February 2025 for studies on surgical management of ELBW neonates with EA/TEF (PROSPERO CRD42025636228). Fatal chromosomal abnormalities were excluded. Demographics, comorbidities, surgical techniques, and complications were analyzed descriptively. Risk of bias was assessed. RESULTS: Eleven publications (five case reports and six case series) comprising 30 patients (Gross type B/C = 1/29) met the eligibility criteria. Mean gestational age was 28.1 (23-34) weeks, and mean birth weight was 760.4 (422-995) g. Twelve primary repairs (PR) and 18 delayed primary repairs (DPR) were performed, including staged repair (n = 11), lower esophageal banding (n = 4), and other techniques (n = 3). Postoperatively, four anastomotic leaks were managed conservatively, six strictures and one recurrent TEF required endoscopic intervention, three fundoplications and two aortopexies were reported (follow-up: 1-198 months, n = 19). Overall mortality was 30% (PR: 8.3%; DPR: 44.4%). Mortality was 60% among neonates with major congenital heart defects (CHD) and 40% among those with VACTERL association. EA/TEF-related complications contributed to 33.3% of deaths. CONCLUSIONS: Mortality in this cohort remains high, particularly with major CHD, and is largely unrelated to EA/TEF-specific complications. In selected cases, PR appears feasible as an alternative to DPR, although conclusions are limited by the small sample size and heterogeneous studies.

Humans

Development and Validation of a Predictive Model for Identification of Cognitive Impairment Risk in Older Adults with Subjective Cognitive Decline:A Longitudinal Study.

BACKGROUND: Subjective cognitive decline (SCD) is a transitional state between objective cognitive impairment and cognitively intact mental status, providing a critical window for implementing preventive interventions to delay objective cognitive decline. AIMS: We aimed to develop a predictive model for SCD progression in older adults with mild cognitive impairment (MCI). This model will facilitate the identification of risk factors and establishment of targeted interventions for community-based SCD management. METHODS: Data from the China Health and Retirement Longitudinal Study (CHARLS) was utilized in this study, extracting 18 indicators. Potential predictors selected through univariate Cox regression and LASSO regression analyses were sequentially incorporated into a multivariable Cox regression model. A nomogram was constructed to establish a predictive model. Model validation encompassed Area Under Curve (AUC) metrics for discriminative capacity, complemented by quantitative assessments using calibration curve analysis for precision verification and decision curve analysis (DCA) for clinical utility evaluation. RESULTS: A total of 1099 older adults with SCD were included in the final analysis, of whom 114 (10.3%) developed MCI. Multivariable Cox regression identified residence, marital status, educational level, social participation, gait speed, and baseline cognitive function. The model demonstrated time-dependent AUC values of 0.885, 0.830, 0.839, and 0.836 in the training set when evaluating discriminative capacity at 2-, 4-, 7-, and 9-year, respectively. The predictive model showed excellent predictive ability according to AUC, calibration curve, and DCA. CONCLUSIONS: A predictive model was created to estimate the risk of developing MCI in older individuals with SCD, offering clinician-actionable intervention benchmarks for preventive care.

Humans

Predictive Models for Hypoglycemia Risk in Haemodialysis Patients With Diabetic Kidney Disease: Systematic Review and Meta-Analysis.

AIM: To provide evidence for selecting and developing reliable clinical assessment tools for hypoglycemia in diabetic kidney disease patients during haemodialysis. DESIGN: Review. METHODS: Systematic searches were performed in 9 Chinese and English databases to collect literature regarding the development of hypoglycemia risk prediction models in haemodialysis patients with diabetic kidney disease. Two reviewers independently performed literature screening, data extraction, risk-of-bias assessment, and applicability evaluation. The Prediction Model Risk of Bias Assessment Tool was used to assess the risk of bias and applicability of the included studies. Meta-analysis was conducted using R software. DATA SOURCES: CNKI, Wanfang, VIP, CBM, PubMed, Cochrane Library, EMbase, Web of Science, and CINAHL. The search period covered from the establishment date of each database to December 2025. RESULTS: Six studies, comprising six prediction models, were included. Two studies performed internal validation, and three conducted external validation. All models reported the area under the curve, ranging from 0.813 to 0.866, and calibration measures. Four studies were rated as having a high risk of bias, while all six demonstrated good overall applicability. The meta-analysis showed that the pooled AUC value of the six studies was 0.846 (95% CI: 0.823-0.867). CONCLUSION: Research on hypoglycemia risk prediction models in haemodialysis patients with diabetic kidney disease remains in the developmental stage. Although the included prediction models exhibited satisfactory apparent discriminatory ability and clinical applicability, most of the original studies suffered from a high risk of bias and lacked adequate validation. The true predictive performance and clinical application value of these models remain to be further verified. Accordingly, routine and unconditional clinical application is not recommended at this stage. Future studies should include more high-quality, multicenter external validation and develop models with high generalizability, favourable clinical applicability, and robust predictive performance to facilitate early identification of hypoglycemia risk in this population. IMPACT: This study systematically evaluated the hypoglycemia risk prediction models for diabetic kidney disease patients during haemodialysis, and the research on hypoglycemia risk prediction models for maintenance haemodialysis patients during dialysis is still in the development stage. This study provides a reference for clinical medical staff to select or develop hypoglycemia risk prediction and assessment tools for diabetic kidney disease patients during haemodialysis. REPORTING METHOD: This study was conducted in accordance with the relevant guidelines of the EQUATOR Network and followed the TRIPOD-SRMA Checklist. PATIENT OR PUBLIC CONTRIBUTION: No patient or public contribution. TRIAL REGISTRATION: PROSPERO: CRD420251243352.

Humans

Artificial intelligence in treatment prediction for skeletal Class III malocclusion: A systematic review.

In skeletal Class III patients, treatment options range from orthodontics to orthognathic surgery. Choosing the optimal approach requires a comprehensive clinical evaluation, which may be supported by AI tools. The aim of this study was to assess the performance of AI models in predicting the need for orthognathic surgery and in identifying predictors influencing treatment decisions. A PRISMA-guided electronic database search (PubMed, Web of Science; 2009-2024; English/French) was performed to identify studies using machine learning (ML) or deep learning (DL) on cephalometric and clinical data. After screening and assessment for eligibility, 15 studies were critically appraised. Model performance was summarized using accuracy, sensitivity, specificity, and the area under the curve (AUC). ML algorithms (particularly Random Forest and XGBoost) and DL models (ResNet-based convolutional neural networks (CNNs)) achieved high accuracy for predicting surgical need. Frequently selected predictors included Wits appraisal, ANB angle, the maxillomandibular ratio (Mx/Md), overjet, and the divergence of the lower gonial angle. AI methods show promise for assisting treatment decisions in Class III malocclusion, with Random Forest and XGBoost performing well on tabular cephalometric data and CNNs on imaging. Larger, multicentre datasets and external validation are needed to improve reliability, address bias, and support clinical implementation.

Humans

Electro-clinical efficacy and safety of midazolam in neonatal seizures: a systematic review with individual level exploratory analysis of gestational age-related treatment response.

UNLABELLED: Neonatal seizures are the most common neurological emergency during the neonatal period and are associated with increased mortality and adverse neurodevelopmental outcomes. Despite current recommendations supporting phenobarbital as first-line therapy, seizure control remains suboptimal in a large proportion of neonates, prompting the use of second-line antiseizure medications. Midazolam is increasingly administered in refractory neonatal seizures but evidence regarding its electro-clinical efficacy and safety remains limited and heterogeneous. To systematically review the available evidence on the electro-clinical efficacy and safety of midazolam in neonatal seizures and to perform an exploratory individual-level analysis investigating the association between gestational age and treatment response. A systematic review was conducted according to PRISMA 2020 guidelines. Studies including neonates with EEG- or aEEG-confirmed seizures treated with midazolam were included. Binary logistic regression was performed to assess the individual-level association between gestational age and treatment response. Eleven studies involving 146 neonates treated with midazolam were included. Electro-clinical response was observed in 101/146 neonates (69.2%), while seizure cessation was achieved in 61/146 neonates (41.8%). In an exploratory complete-case logistic regression analysis, higher gestational age appeared to be associated with a greater probability of electro-clinical response. The predicted probability curve crossed the 50% response probability at approximately 36.5 weeks of gestation. Hypotension was the most frequently reported adverse event, while respiratory depression, sedation-related effects, and transient EEG/aEEG suppression were reported less frequently. CONCLUSIONS: Midazolam may have a role as an add-on antiseizure medication in neonatal seizures, particularly in refractory cases. However, the evidence remains limited by heterogeneity in study design, EEG monitoring strategies, outcome definitions, and incomplete individual-level data. The observed association between gestational age and response is hypothesis-generating and requires prospective validation. WHAT IS KNOWN: • Phenobarbital often provides incomplete seizure control in neonates, making second-line antiseizure therapies necessary in refractory cases. • Evidence supporting midazolam for neonatal seizures remains limited and heterogeneous. WHAT IS NEW: • This systematic review summarizes the electro-clinical efficacy and safety of midazolam and includes an exploratory patient-level analysis suggesting that higher gestational age may be associated with improved treatment response. • These findings support further prospective studies on developmental determinants of response to GABAergic therapy.

Humans

Predicting ACL injury risk in athletes: A systematic review of machine learning-based models.

BACKGROUND: Early ACL injury risk identification in athletes is essential. This systematic review examines machine learning (ML) models for predicting ACL injuries, evaluating their methodological quality, performance, and reliability. METHOD: A comprehensive electronic search was conducted across PubMed, Scopus, Web of Science, and IEEE Xplore databases, supplemented by Google Scholar for grey literature, covering articles published between January 1, 2015, and August 30, 2025. Eligible studies were appraised using the Prediction Model Study Risk of Bias Assessment Tool (PROBAST) for methodological quality and risk of bias, and the Transparent Reporting of a Multivariable Prediction Model for Individual Prognosis or Diagnosis (TRIPOD) guidelines for quality of evidence. RESULTS: Ten studies were included. PROBAST showed eight studies had moderate risk of bias and two low risk. TRIPOD found only two studies met quality criteria. ML models included logistic regression (n = 5), support vector machines (n = 4), k-nearest neighbor (n = 3), decision trees (n = 3), random forests (n = 5), neural networks (n = 2), linear discriminant analysis (n = 1), and pre-trained CNNs (n = 1). AUC ranged from 0.63 to 0.98. Accuracy (reported in six studies) ranged from 26% to 95%; however, these values should be interpreted with caution due to the absence of confidence intervals, lack of class imbalance handling, and limited external validation across studies. Tree-based ensemble methods such as random forest achieved competitive accuracy (74-86%), while SVM, a non-ensemble classifier, reported accuracy ranging from 71% to 95%; however, the highest values were obtained in studies with notably small sample sizes (n = 12 to n = 39), raising concerns about overfitting and generalizability. CONCLUSION: Current ML algorithms show promise for identifying athletes at high ACL injury risk and detecting relevant risk factors. Although study quality was generally satisfactory, future research should prioritize external validation and model interpretability to support clinical translation.

Humans

Artificial intelligence-supported double reading in European population breast cancer screening: A systematic review and meta-analysis of prospective programs.

BACKGROUND: Most European population mammography screening programs rely on double reading with arbitration, a model that delivers mortality benefit but is increasingly challenged by radiologist workload, variable specificity, and interval cancers. Artificial intelligence (AI) is being evaluated to support or optimize these established European screening pathways. PURPOSE: To synthesize prospective or program-embedded evaluations of AI conducted within European-style population screening programs and to estimate exploratory program-level absolute risk differences (RDs) per 1000 examinations for cancer detection rate (CDR) and recall. MATERIALS AND METHODS: We performed a prespecified, focused evidence synthesis of three large studies embedded within routine population screening programs operating under European-relevant workflows: MASAI (randomized AI-supported risk triage within a national program), ScreenTrustCAD (prospective paired-reader evaluation with AI as an independent reader in a double-reading framework), and PRAIM (nationwide decision-referral implementation). Outcomes were harmonized as AI-control RDs per 1000 examinations. Random-effects pooling used Hartung-Knapp-Sidik-Jonkman models. For the paired-reader design, sensitivity analyses applied a Kish effective sample-size approach across plausible within-examination correlations (ρ = 0.3-0.8). Positive predictive value (PPV) and workflow/time outcomes were summarized descriptively. RESULTS: Across 597,419 examinations, the pooled CDR RD was +0.9 per 1000 (95% CI -0.0 to +1.8; I2 ≈ 12%), consistent with a modest directional increase with borderline statistical uncertainty. The pooled recall RD was -0.6 per 1000 (95% CI -3.1 to +2.1; I2 ≈ 41-43%), indicating no consistent recall increase across screening programs. Where reported, PPV was higher with AI-supported screening. Efficiency signals included 44.3% fewer total readings in MASAI and shorter reading times for AI-normal examinations in PRAIM; in PRAIM, a program-level safety-net mechanism recovered 204 cancers that would otherwise have been missed. CONCLUSION: In European population screening programs characterized by double reading and arbitration, prospective program-embedded evidence suggests that AI integration may yield a small absolute increase in cancer detection (≈1/1000) without a consistent increase in recall, alongside improved PPV and efficiency signals. These findings suggestAI primarily as a complementary reader within European screening workflows, with implementation requiring explicit quality assurance and monitoring of interval cancers and stage distribution.

Humans

Reinforcement learning-based dynamic ensemble for missense variant effect prediction and tiered prioritization of VUS.

BACKGROUND: Accurate classification of missense variants remains a challenging task despite major advances in genomics. Numerous computational models have been developed to assist in variant classification, but often require repeated integration and benchmarking efforts. Ensemble methods have been proposed to overcome the limitations of single predictors, but mostly rely on fixed, predefined weights that constrain their ability to capture interactions among predictive signals. METHODS: We present GenixRL, a dynamic ensemble framework that reformulates model fusion as a reinforcement learning optimization problem. GenixRL uses a Q-learning agent to learn a policy that dynamically weights the probabilistic outputs of complementary predictors, including BayesDel (addAF and noAF), ClinPred, and MetaRNN. Replacing static weighting with policy learning allows GenixRL to adaptively identify optimal weightings and substantially improve classification accuracy. RESULTS: In benchmark evaluation against 25 state-of-the-art predictors, GenixRL achieved an AUROC of 0.9644 on an independent ClinVar dataset. On saturation genome editing assays for BRCA1 and BRCA2, GenixRL achieved the best performance and ranked highest on 14 of 17 clinically significant genes in a zero-shot evaluation. Applied to uncertain and conflicting ClinVar variants, GenixRL enabled tiered, evidence-based prioritization of hundreds of thousands of variants as likely pathogenic or pathogenic with high confidence, supported by orthogonal population evidence from gnomAD. CONCLUSION: GenixRL advances pathogenicity prediction for missense variants and provides an adaptive ensemble that sorts variants of uncertain significance into tiered candidates for expert curation and functional validation.

Mutation, Missense

Gastrointestinal digestion governs insect protein hydrolysis and predicted bioactive peptide release: Species-dependent implications for functional food applications.

This study investigates the digestion of insect proteins and the release of predicted bioactive peptides during human gastrointestinal digestion. Using the Infogest in vitro model, mealworm, cricket, and black soldier fly larvae (BSFL) proteins were digested and analyzed through discovery proteomics and bioinformatics to identify predicted bioactive peptides. Sequential windowed acquisition of all theoretical fragment ion mass spectra (SWATH-MS) quantified insect proteins including predicted bioactive peptide precursor proteins, the precursors of predicted bioactive peptides. Results indicated that gastrointestinal digestion strongly influences peptide release, with the gastric phase exhibiting a richer predicted bioactive peptide profile than the small intestinal phase. Many predicted bioactive peptides were rapidly hydrolysed under small intestine conditions, which may lead to reduced stability or diminished activity in vivo, potentially explaining why certain peptides show strong bioactivity in vitro but limited effects in vivo. Additionally, predicted bioactive peptide release varied by insect species, influenced by genetic factors and peptide abundance. These findings highlight the importance of species selection and consideration of proteolytic digestion patterns in optimizing insect-derived bioactive peptides for functional foods and nutraceutical applications.

Animals

Whole-Genome Deep Learning Predicts Chemotherapy Response in Colorectal Cancer.

Chemotherapy response in colorectal cancer (CRC) exhibits significant heterogeneity, with current clinical predictors failing to capture complex genomic determinants of resistance. We developed a hybrid deep learning framework integrating convolutional neural networks (CNNs) and bidirectional long short-term memory (BiLSTM) networks to analyze whole-genome somatic mutations, evolutionary conservation, chromatin accessibility, and 3D genome architecture in 2,546 TCGA patients. An attention mechanism identified predictive genomic regions. The model achieved an AUC of 0.92 (95% CI: 0.89-0.94) in cross-validation and 0.88 (95% CI: 0.85-0.91) in independent validation, outperforming clinical models (&#x394;AUC = +0.18, p < 0.001). Key predictors included non-coding variants in TP53, KRAS, and PIK3CA regulatory regions. Triple-positive patients (mutations in all 3 regions) had significantly worse progression-free survival (HR = 4.7, p < 0.001). Our framework enables accurate chemotherapy response prediction and reveals novel non-coding resistance mechanisms, advancing precision oncology in CRC.

Humans

Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.

BACKGROUND: Predictive modelling for cancer risk, treatment-related complications, and survival is central to precision oncology. Conventional logistic regression (LR) and Cox proportional hazards (CoxPH) regression remain widely used but are limited when modelling nonlinear interactions, high-dimensional imaging features, and multimodal clinical-metabolic predictors. Artificial intelligence (AI) and machine learning (ML) methods offer expanded capability through automated feature extraction, ensemble learning, and flexible survival modelling, but the evidence on when AI/ML adds value over conventional models across cancer sites and predictive tasks remains fragmented. OBJECTIVE: To systematically evaluate the methodological performance, validation strategies, and translational limitations of AI/ML models compared with conventional statistical models in published predictive-modelling studies for breast, colorectal, or pancreatic cancer. METHODS: PubMed, Scopus, and Web of Science were searched for studies published between January 2019 and March 2025. Two reviewers independently conducted title-and-abstract screening, full-text eligibility assessment, and PROBAST risk-of-bias assessment. Sixty-five studies (n&#xa0;=&#xa0;907,567 participants) were narratively synthesised by cancer site, predictive task, model family, comparator, validation strategy, predictor modality, and calibration or explainability reporting. RESULTS: The 65 studies comprised breast cancer (n&#xa0;=&#xa0;35), colorectal cancer (n&#xa0;=&#xa0;21), and pancreatic cancer (n&#xa0;=&#xa0;9). AI/ML superiority over LR and CoxPH was task- and data-dependent. CNN- and U-Net-based models predominated in imaging and body-composition tasks, tree-based ensembles consistently outperformed LR for tabular perioperative complication prediction, and CoxPH remained competitive, and in the largest pancreatic risk study, superior to XGBoost (C-index 0.802 vs 0.723) in well-structured datasets. PROBAST analysis-domain risk was moderate in 54 of 65 studies (83%), driven by limited external validation, sparse calibration reporting (11/65), and few decision-curve analyses (7/65). CONCLUSION: AI/ML adds the most methodological value in imaging-derived feature extraction and nonlinear perioperative prediction, while conventional regression remains preferable in large, structured datasets with linear predictors. Clinical translation requires standardised body-composition definitions, external validation, calibration assessment, decision-curve analysis, and explainability, in line with TRIPOD+AI and CLAIM standards.

Humans

Predictive value of anal sphincter electromyography for sacral neuromodulation test-phase outcomes.

BACKGROUND: Sacral neuromodulation (SNM) is an established therapy for refractory pelvic organ dysfunction. Anal sphincter electromyography (EMG) is commonly used preoperatively to assess sacral and peripheral nerve integrity. However, the prognostic significance of chronic neurogenic EMG changes for SNM outcomes remains unclear. OBJECTIVE: To evaluate whether chronic neurogenic changes on preoperative anal sphincter EMG predict the outcome of the SNM test phase. METHODS: We retrospectively analysed 62 consecutive patients with bladder and/or bowel dysfunction or pelvic pain who were candidates for SNM treatment and who underwent preoperative anal sphincter EMG. EMG findings were classified as normal or showing chronic neurogenic changes. SNM test-phase success was defined as a &#x2265;50% improvement of symptoms at 24&#xa0;days. Outcomes were compared between EMG groups. RESULTS: Of the 62 patients (49 women, 13 men), 30 (48%) had normal EMG findings and 32 (52%) showed chronic neurogenic changes. Overall, the SNM test phase was successful in 47 patients (76%). Success rates were similar in patients with normal EMG (72%) and neurogenic EMG changes (79%), with no statistically significant difference (p&#xa0;=&#xa0;0.878). Sex-stratified analyses revealed no significant association between EMG findings and test-phase success in women or men. CONCLUSIONS: Chronic neurogenic changes on anal sphincter EMG do not predict SNM test-phase outcomes. These findings suggest that abnormal sphincter EMG results should not be used as a standalone criterion to exclude patients from SNM therapy. SIGNIFICANCE: Signs of neurogenic damage on anal sphincter EMG are not an indicator of reduced neuromodulatory capacity or diminished clinical response to SNM.

Humans