Search PubMedSearch

SEARCH · Search PubMed

Results for “Learning Curve”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

57 recordsLinked to original sources

Distal versus proximal radial access for diagnostic cerebral angiography: comparative outcomes and learning curve analysis.

BACKGROUND AND PURPOSE: Distal transradial access (dTRA) is an alternative to proximal transradial access (pTRA) for neuroangiography, but comparative real-world data and evidence on its early learning curve remain limited. We compared procedural performance and access-site complications between dTRA and pTRA and evaluated the early learning curve of dTRA. METHODS: We retrospectively analyzed 470 diagnostic cerebral angiography procedures, representing 421 unique patients, performed via radial access at a single center between January 2025 and February 2026, including 237 dTRA and 233 pTRA procedures. Baseline characteristics, including age, sex, body mass index (BMI) category, aortic arch type, and antiplatelet/anticoagulant use, procedural performance, and clinically assessed access-site events were compared between groups. Radial artery occlusion (RAO) was assessed by postoperative bedside pulse examination and confirmed with Doppler ultrasound when clinical findings were uncertain. Multivariable logistic regression was used to evaluate predictors of RAO, persistent bleeding or repeated compression, hand edema, and a composite access-site event endpoint. Because repeated procedures occurred in a subset of patients and event counts were limited, first-procedure sensitivity analysis and analyses of infrequent outcomes were interpreted cautiously. The dTRA learning process was assessed in the first 100 dTRA cases performed by a single operator using multivariable regression, cumulative sum (CUSUM) analysis, segmented trend analysis, and phase-based comparisons. RESULTS: Baseline characteristics were comparable between groups, including age, male sex, BMI category, aortic arch type, and antiplatelet/anticoagulant use. Compared with pTRA, dTRA was associated with more puncture attempts (3.0 [2.0-4.0] vs 2.0 [1.0-3.0], P&#xa0;<&#xa0;0.001), longer puncture time (2.0 [1.0-5.0] vs 2.0 [1.0-3.0] min, P&#xa0;=&#xa0;0.003), lower first-pass success (19.4% vs 35.2%, P&#xa0;<&#xa0;0.001), and a higher crossover rate (11.4% vs 6.0%, P&#xa0;=&#xa0;0.037). However, dTRA was associated with a lower clinically assessed RAO rate (2.5% vs 7.7%, P&#xa0;=&#xa0;0.011). On multivariable analysis, pTRA was independently associated with higher odds of RAO (OR 3.27, 95% CI 1.26-8.49, P&#xa0;=&#xa0;0.015) and the composite access-site event endpoint (OR 3.12, 95% CI 1.55-6.28, P&#xa0;=&#xa0;0.001). Similar findings were observed in a sensitivity analysis restricted to the first procedure per patient. In the first 100 dTRA cases, cumulative dTRA experience was independently associated with shorter total procedure time (beta&#xa0;=&#xa0;-0.074&#xa0;min/case, P&#xa0;=&#xa0;0.009), while CUSUM and moving-average analyses suggested that the major learning effect occurred within approximately the first 10-15 cases. CONCLUSIONS: In this retrospective single-operator cohort, dTRA was associated with lower clinically assessed RAO than pTRA despite greater access difficulty. The early learning effect was mainly reflected in shorter total procedure time. These findings support the feasibility of dTRA but should be interpreted cautiously given the study's observational design and limited anatomical data.

Humans

Enhanced fracture detection on radiographs with AI assistance for clinicians: a systematic review and meta-analysis.

BACKGROUND: Emergency radiographic interpretation for fractures is prone to missed or misdiagnoses. Artificial intelligence (AI) is expected to become a powerful tool to assist clinicians in fracture detection. PURPOSE: A systematic review and meta-analysis was performed to assess whether AI improves clinicians' ability to detect fractures on radiographs. MATERIALS AND METHODS: A literature search was conducted in PubMed, Web of Science, and Cochrane Library for studies published between January 1, 2010, and October 10, 2025. A meta-analysis of diagnostic accuracy studies was performed using a Summary Receiver Operating Characteristic (SROC) curve. The quality of included studies was assessed using the Quality Assessment of Diagnostic Accuracy Studies 2 (QUADAS-2) tool. Subgroup analysis and meta-regression were conducted to explore potential sources of heterogeneity. RESULTS: A total of 26 studies were included . The pooled sensitivity of clinicians increased from 77% (95% CI: 72-81) to 87% (95% CI: 83-90) with AI assistance, while the pooled specificity improved from 88% (95% CI: 85-90) to 92% (95% CI: 89-94). The corresponding AUC values were 0.90 (95% CI: 0.87-0.92) before and 0.95 (95% CI: 0.93-0.97) after AI assistance. Eight studies were rated as high risk of bias. Subgroup analysis and meta-regression identified potential sources of heterogeneity, including fracture location, AI model type, high risk of bias, and reference standards. CONCLUSION: AI assistance significantly improves clinicians' diagnostic performance in detecting fractures on radiographs for extremity and trunk fractures.

Humans

Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.

BACKGROUND: Predictive modelling for cancer risk, treatment-related complications, and survival is central to precision oncology. Conventional logistic regression (LR) and Cox proportional hazards (CoxPH) regression remain widely used but are limited when modelling nonlinear interactions, high-dimensional imaging features, and multimodal clinical-metabolic predictors. Artificial intelligence (AI) and machine learning (ML) methods offer expanded capability through automated feature extraction, ensemble learning, and flexible survival modelling, but the evidence on when AI/ML adds value over conventional models across cancer sites and predictive tasks remains fragmented. OBJECTIVE: To systematically evaluate the methodological performance, validation strategies, and translational limitations of AI/ML models compared with conventional statistical models in published predictive-modelling studies for breast, colorectal, or pancreatic cancer. METHODS: PubMed, Scopus, and Web of Science were searched for studies published between January 2019 and March 2025. Two reviewers independently conducted title-and-abstract screening, full-text eligibility assessment, and PROBAST risk-of-bias assessment. Sixty-five studies (n&#xa0;=&#xa0;907,567 participants) were narratively synthesised by cancer site, predictive task, model family, comparator, validation strategy, predictor modality, and calibration or explainability reporting. RESULTS: The 65 studies comprised breast cancer (n&#xa0;=&#xa0;35), colorectal cancer (n&#xa0;=&#xa0;21), and pancreatic cancer (n&#xa0;=&#xa0;9). AI/ML superiority over LR and CoxPH was task- and data-dependent. CNN- and U-Net-based models predominated in imaging and body-composition tasks, tree-based ensembles consistently outperformed LR for tabular perioperative complication prediction, and CoxPH remained competitive, and in the largest pancreatic risk study, superior to XGBoost (C-index 0.802 vs 0.723) in well-structured datasets. PROBAST analysis-domain risk was moderate in 54 of 65 studies (83%), driven by limited external validation, sparse calibration reporting (11/65), and few decision-curve analyses (7/65). CONCLUSION: AI/ML adds the most methodological value in imaging-derived feature extraction and nonlinear perioperative prediction, while conventional regression remains preferable in large, structured datasets with linear predictors. Clinical translation requires standardised body-composition definitions, external validation, calibration assessment, decision-curve analysis, and explainability, in line with TRIPOD+AI and CLAIM standards.

Humans

Development and validation of a comprehensive prognostic model for 28-day ICU mortality in non-traumatic subarachnoid hemorrhage: an analysis based on the MIMIC-IV database.

BACKGROUND: Due to the complex pathophysiology of non-traumatic subarachnoid hemorrhage (SAH), accurate risk prediction remains a challenge. Our aim is to develop and validate a comprehensive prognostic model that integrates demographic characteristics, vital signs, laboratory parameters, and more, to provide clinical decision-making support in real-world practice. METHODS: We conducted a retrospective cohort study of 785 Non-traumatic subarachnoid hemorrhage patients. The cohort was randomly divided into a training set (n&#xa0;=&#xa0;549) and a validation set (n&#xa0;=&#xa0;236). Feature selection was performed using LASSO regression, followed by backward stepwise Cox regression for optimization. A nomogram was constructed based on independent predictive factors, and model performance was assessed using discrimination, calibration, and decision curve analysis. To prevent immortal-time bias, all predictors were anchored to a fixed early (first-24-hour) measurement window, treatment variables were modelled as binary indicators rather than cumulative exposures, and a five-model sensitivity analysis with baseline-severity adjustment was performed. RESULTS: The development of our model followed a systematic approach: first, 15 potential predictive factors were selected via LASSO regression, which were then refined to 12 independent predictors using backward stepwise Cox regression. The final predictive factors included: Ventilation, AHT, Nimodipine 60&#xa0;mg, Age, SAPS.II, Input amount, Calcium total, Platelet count, White blood cells, Anion gap, pH, and Chloride. The integrated model demonstrated excellent predictive ability for 7-day, 14-day, and 21-day mortality in both the training set (AUC: 0.972, 0.934, 0.898) and the validation set (AUC: 0.968, 0.948, 0.911). Calibration curves and decision curve analysis confirmed the model's reliability and clinical utility across different time points. We constructed a nomogram for individualized risk prediction. Univariate Kaplan-Meier survival analysis demonstrated significant stratification of survival outcomes by each predictor, while restricted cubic spline analysis revealed non-linear relationships between continuous variables and mortality risk. Random survival forest analysis identified the top three predictive factors (Nimodipine 60&#xa0;mg, Ventilation, AHT) and compared them with our full 12-variable model, confirming superior performance of the integrated model at all time points. At the 28-day primary endpoint, the model achieved a time-dependent AUC of 0.898 (training) and 0.904 (validation); after restricting predictors to the early baseline window, the leakage-controlled model retained good discrimination (validation C-index 0.803). CONCLUSIONS: Our ICU 28-day mortality prognosis model demonstrated robust performance in predicting ICU 28-day mortality in non-traumatic subarachnoid hemorrhage. The model, through the nomogram, provides individualized risk assessment, aiding clinical decision-making and patient stratification.

Humans

Integrative machine learning and transcriptomic analysis reveals molecular mechanisms underlying low survival rate in larval Chinese Bahaba (Bahaba taipingensis).

Chinese Bahaba (Bahaba taipingensis) is a Class I protected marine fish endemic to China. Low larvae survival during artificial breeding severely hinder population recovery. To investigate the molecular mechanism of high mortality in larval fish, this study performed RNA-seq on liver from naturally deceased (ND) and mass-dead (MD) individuals, combined with least absolute shrinkage and selection operator (LASSO) regression and random forest (RF) algorithms to screen for core signature genes. A total of 873 differentially expressed genes (DEGs) were identified, including 112 upregulated and 761 downregulated genes. GO and KEGG enrichment analyses revealed significant enrichment in amino acid metabolism disorders, one&#x2011;carbon folate pool impairment, PPAR signaling abnormalities, ECM-receptor interaction, focal adhesion pathway, indicating widespread metabolic suppression accompanied by extracellular matrix remodeling and signaling disturbances in the livers of MD fish. MAD pre-filtering combined with dual machine learning algorithms yielded 18 robust core signature genes, among which SLC38A4, MMP1, FADD, FKBP5, and APOB were consistently identified as high-frequency core genes by both algorithms. SLC38A4 exhibited the highest importance score in the RF model and was significantly downregulated, making it the primary molecule distinguishing ND from MD phenotypes. ROC curve analysis showed that both models achieved an AUC of 1.000 (95% CI lower bound: 0.610), confirming the precise discriminatory ability of the core genes. GSEA further demonstrated significant enrichment of this core gene set in ND samples. This study provides the first systematic elucidation of the molecular mechanisms underlying liver dysfunction in low survival rate B. taipingensis, characterized by amino acid transport impairment, metabolic reprogramming, and structural remodeling, offering theoretical foundations for health assessment, early mortality risk warning, and artificial breeding conservation of this species.

Animals

Could the preoperative urethral curve be used to predict immediate urinary continence following Retzius-sparing robot-assisted radical prostatectomy? A retrospective multi-center study.

PURPOSE: Immediate urinary continence (UC) recovery following Retzius-sparing robot-assisted radical prostatectomy (RS-RARP) remains highly variable, highlighting the need for reliable preoperative prediction. We aimed to develop and validate models to identify patients likely to achieve immediate UC recovery following RS-RARP. MATERIALS AND METHODS: A total of 580 prostate cancer patients who underwent RS-RARP from four medical centers were assigned to a training set (n=348), an internal validation set (n=103) and an external validation set (n=129). Independent predictors were identified through univariate analysis and LASSO regression. A nomogram was constructed using multivariate logistic regression. Its performance was evaluated with receiver operating characteristic (ROC) curve, calibration curves, and decision curve analysis. RESULTS: Immediate UC recovery was observed in 84.5% (294/348) of patients in the training cohort, 80.6% (83/103) in the internal validation cohort, and 81.4% (105/129) in the external validation cohort, respectively. Multivariate analysis identified membranous urethral length (MUL) (OR=1.23, P=0.029) and urethral curvature (OR=2.84, P<0.001) as independent predictors, while prostate volume (PV) (OR=0.84, P <0.001) as a protective factor. The nomogram integrating MUL, PV, and urethral curvature demonstrated superior predictive accuracy, with an AUC of 0.87 (95% CI, 0.83-0.91) in the training cohort. The bootstrap-corrected calibration slope was 0.96, and the Brier score was 0.08.&#xa0;Calibration curves and decision curve analysis confirmed the predictive accuracy and clinical utility of the nomogram. CONCLUSIONS: Our study introduces a novel quantitative method for assessing urethral curvature. The mpMRI-based model, integrating urethral curvature and prostate spatial configuration, offers enhanced predictive accuracy for postoperative immediate UC recovery.

Humans

Advancing nursing education through social and emotional learning: A systematic review guided by the Collaborative for Academic, Social, and Emotional Learning framework.

BACKGROUND: With Generation Z entering the nursing workforce in growing numbers, strengthening social and emotional learning is critical for academic success, professional adaptation, and safe practice. However, the existing evidence remains fragmented because of varied interventions and inconsistent approaches. OBJECTIVES: This systematic review examined (1) the social and emotional learning essential for nursing students and nurses within the Collaborative for Academic, Social, and Emotional Learning framework, (2) their impact on educational and clinical outcomes, and (3) implications for advancing nursing education and practice. METHODS: Following Joanna Briggs Institute methodology and Preferred Reporting Items for Systematic Reviews and Meta-Analyses guidelines, five international (PubMed, EMBASE, CINAHL, PsycINFO, Cochrane) and three Korean (RISS, KoreaMed, KMBASE) databases were searched up to June 2025. Eighteen studies involving 2,952 participants met the inclusion criteria, including quasi-experimental quantitative studies, descriptive quantitative studies, qualitative studies, and mixed-methods studies. The methodological quality of the included studies was appraised using the Mixed Methods Appraisal Tool. RESULTS: Within the Collaborative for Academic, Social, and Emotional Learning framework, relationship skills and self-management were the most frequently studied competencies, emphasizing teamwork, communication, and stress regulation. Self-awareness and social awareness were underexplored, despite their importance in empathy, resilience, and reflective practice. Responsible decision-making was the least studied competency, despite its importance in ethical reasoning. Social and emotional learning was consistently associated with enhanced adaptation, communication, leadership, relationships, and clinical performance. Effective strategies included blended learning, simulation, reflective activities, and mentorship, which are aligned with Generation Z's learning preferences. CONCLUSION: Although social and emotional learning integration is associated with improvements in educational and clinical outcomes in nursing, current research has largely centered on relational and stress-related competencies while underrepresenting responsible decision-making. To cultivate reflective, empathetic, and ethically grounded nurses, curricula should integrate social and emotional learning through a balanced and structured approach. REGISTRATION: This study was registered on PROSPERO (ID: CRD420251005683).

Humans

Analysis of deep learning techniques in computer-aided diagnosis for meniscus injuries: a systematic literature review.

Meniscus informatics is a growing subject of study in the healthcare industry. One of the major hindrances to the healthcare system's transformation is obtaining knowledge and meaningful information from complicated, high-dimensional and diverse sources. Modern biomedical research, for instance, has seen an increase in the use of complex, dissimilar, poorly documented, and generally unstructured electronic health records, imaging, sensor data and text, even after many current techniques have been used to extract more robust and useful elements from the data for analysis. New efficient standards for building end-to-end learning models from complex data are therefore needed. Therefore, the current study aims to examine the most recent research on the use of deep learning techniques for diagnosing meniscus tears and recommend creating comprehensive and meaningful interpretable structures that might benefit the healthcare industry. We also draw attention to shortcomings and the need for better technique development, and we provide new perspectives about this exciting new development in the field.

Humans

A systematic review of human avoidance learning: Cognition, computation, and methods.

Avoidance behaviour is fundamental for survival but can become maladaptive in clinical conditions. A large body of literature has accumulated on the dynamics of human avoidance learning. However, current theories and overviews do not provide an exhaustive account of this evidence. In this systematic review, we identify N = 116 studies on human avoidance learning. We analyse these studies with the goal of distilling robust empirical phenomena as a basis for theory-building, and examine their diagnostic value in differentiating between competing theories. We find that the evidence is difficult to reconcile with foundational two-factor and classical safety-signal accounts, and most strongly supports expectancy- and inference-based views, in which avoidance responses are selected with respect to represented consequences. At the same time, no current framework provides a complete account of the evidence: several findings point to an additional role for operant valuation, Pavlovian influences, and contextual or latent-state control over the expression of avoidance. Methodologically, we observe that the problem setting in the most common experimental paradigms is radically simpler than real-world avoidance and therefore unlikely to expose the limits of inferential or reflective mechanisms. Consequently, we argue that paradigms with greater computational demands and more realistic action affordances are required to identify the mechanisms underlying avoidance learning. Collectively, these insights provide a foundation for theoretical refinement, computational modelling, and methodological innovation, with implications for advancing interventions targeting maladaptive avoidance.

Humans

Machine learning-based prediction of unplanned readmission and construction of an online calculator for elderly patients with mild ischemic stroke.

OBJECTIVE: To screen for independent risk factors for unplanned readmission in elderly patients with mild ischemic stroke, and to construct and validate an online risk prediction calculator based on an interpretable machine learning model, thereby providing a promising practical tool for accurate clinical assessment of 30&#x2011;day all&#x2011;cause unplanned readmission risk in this population. METHODS: A prospective cohort study was conducted, including 1050 patients aged&#xa0;&#x2265;&#xa0;60&#xa0;years with mild ischemic stroke admitted between August 2023 and September 2024. Participants were randomly divided into a training set (840 cases) and a test set (210 cases) at a ratio of 8:2. Risk factors were screened by univariate analysis and multivariable Logistic regression. Four machine learning models, namely LightGBM, XGBoost, Random Forest, and K&#x2011;Nearest Neighbors (KNN), were developed and their performance was evaluated using AUC, accuracy, sensitivity, and specificity as metrics. The SHAP framework was used for interpretability analysis, and an online calculator was subsequently developed based on the optimal model. RESULTS: Univariate analysis showed significant differences (P&#xa0;<&#xa0;0.05) in 13 factors including age, smoking, AIP, TyG index, HALP score, etc. Multivariable Logistic regression identified age (OR&#xa0;=&#xa0;9.752), smoking (OR&#xa0;=&#xa0;5.171), AIP (OR&#xa0;=&#xa0;6.691), TyG index (OR&#xa0;=&#xa0;4.393), HALP score (OR&#xa0;=&#xa0;2.831), and&#xa0;&#x2265;&#xa0;2 comorbidities (OR&#xa0;=&#xa0;3.664) as independent risk factors. All four machine learning models demonstrated good predictive performance. Based on a comprehensive evaluation of multiple metrics and computational efficiency, the LightGBM model exhibited the best predictive performance (AUC&#xa0;=&#xa0;0.884, accuracy&#xa0;=&#xa0;0.829, sensitivity&#xa0;=&#xa0;0.812, specificity&#xa0;=&#xa0;0.875). SHAP analysis showed that age, AIP, TyG index, smoking, and HALP score were key predictors. An online calculator developed based on this model enables individualized risk predictions. CONCLUSION: Key risk factors associated with 30&#x2011;day unplanned readmission in elderly patients with mild ischemic stroke were identified. The LightGBM model demonstrated high predictive accuracy, and together with the interpretability analysis and online calculator, offers a practical tool to support clinical risk assessment. However, this tool requires future external validation.

Humans

Diagnostic performance of machine learning models for malignant and non-malignant pleural effusion: Systematic review and meta-analysis.

BACKGROUND: Accurately distinguishing malignant pleural effusion (MPE) from non-malignant pleural effusion is clinically important, but the generalisability and methodological quality of machine-learning (ML) models remain uncertain. METHODS: We searched eight databases to 23 April 2026. Diagnostic performance was pooled using random-effects and Reitsma bivariate models, and study quality was assessed using PROBAST+AI. RESULTS: Forty-two studies were included; 17 contributed to the AUC meta-analysis and 14 to the bivariate analysis. The pooled AUC was 0.90 (95&#xa0;% CI 0.85-0.94; 95&#xa0;% prediction interval 0.62-0.98), with sensitivity of 0.80 (95&#xa0;% CI 0.77-0.83) and specificity of 0.87 (95&#xa0;% CI 0.79-0.92). Only nine studies reported external, temporal or independent validation. Externally validated studies had a lower pooled AUC than studies without external validation (0.83 vs 0.92), with lower specificity observed in the two externally validated studies contributing sensitivity and specificity data. All 42 development assessments had high overall quality concerns, and all 42 model evaluations were judged at high risk of bias. CONCLUSIONS: ML models showed good apparent accuracy for distinguishing MPE from non-MPE, but the evidence was limited by substantial heterogeneity, high risk of bias and scarce external validation. The pooled estimates reflect the average performance of different selected models rather than the expected accuracy of a single clinical test. ML models should be regarded as adjuncts to existing diagnostic pathways until they are confirmed by rigorous multicentre prospective external validation and clinical-impact studies.

Humans

Effects of extended problem-based learning interventions on undergraduate nursing education: A systematic review.

OBJECTIVE: Exploring the effects of long-term PBL (problem-based learning) intervention on undergraduate nursing students. METHODS: The article retrieved literature from CINAHL Complete, Academic Search Complete, Web of Science, PubMed, EMBASE, OVID, and Cochrane Library up to January 2025. Studies had to meet all of these criteria: (1) They used a randomized controlled trial (RCT) and quasi-experimental design. (2) The PBL pedagogy intervention lasted 4&#xa0;weeks or longer. (3) The participants were undergraduate nursing students. (4) They reported primary outcomes. These included critical thinking, problem-solving skills and self-directed learning. Two researchers screened articles, extracted data, and assessed quality independently using blinding. They used Cochrane ROB2 for RCTs and ROBINS-I for quasi-experimental studies to judge bias risk. Meta-analysis was performed using RevMan 5.4 software. For continuous variables, standardized mean difference (SMD) and 95% confidence interval were calculated. Heterogeneity was assessed by I2 statistic. When I2&#xa0;>&#xa0;50%, sensitivity analysis was conducted. The source of heterogeneity was explored by excluding studies one by one. The primary outcomes included standardized critical thinking, problem-solving, and self-directed learning assessment results. RESULTS: A total of 11 randomized controlled trials and quasi-experimental studies were retrieved and included for meta-analysis. The experimental group significantly outperformed the control group in critical thinking, problem-solving, and self-directed learning, with differences being statistically significant (P&#xa0;&#x2264;&#xa0;0.05). However, high heterogeneity was observed. After sensitivity analysis, the heterogeneity was reduced and the results remained statistically significant, indicating that the findings were not solely dependent on the excluded studies.

Problem-Based Learning

Comparative effectiveness of game-based learning modalities in nursing and medical education: a systematic review and Bayesian network meta-analysis.

BACKGROUND: Game-based learning (GBL) is increasingly used in healthcare education, but educators must choose among diverse modalities (e.g., quiz platforms, apps, serious games and metaverse environments). Comparative evidence on which modalities perform best across learning domains (knowledge, attitudes, and practice) remains limited. AIM: To compare the effects of distinct GBL modalities on knowledge, attitudes, and practice outcomes in nursing and medical education and to explore whether comparative effects differ by learner group (pre-licensure students and in-service professionals). DESIGN: PRISMA-NMA-aligned systematic review and Bayesian network meta-analysis. METHODS: We searched eight databases and trial registries through September 2, 2024, for randomized controlled trials comparing GBL with traditional teaching (TT). Outcomes were transformed to a 0-100 scale and analysed as change from baseline in Bayesian consistency models; random-effects models were selected using deviance information criterion (DIC). Risk of bias was assessed using RoB 2. We report mean differences (MDs) with 95% credible intervals (CrIs) versus TT, ranking probabilities, and subgroup NMAs by learner group. RESULTS: Thirty-one RCTs (n&#xa0;=&#xa0;3439) were included; 15 contributed complete data to the network. Risk of bias was low in 15 trials and raised some concerns in 16. The network was modest for knowledge (11 trials) and sparse for attitudes (3) and practice (4). Compared with TT, metaverse-based learning showed improved attitudes (MD 15; 95% CrI 12 to 18), based on a single trial. For knowledge and practice, Kahoot-based quizzes (MD 9.1; 95% CrI -8.9 to 27) and app-based learning (MD 4.6; 95% CrI -4.4 to 14) had the highest estimated mean improvements, but credible intervals were wide and included the null for most comparisons. Subgroup rankings differed by learner group, but several comparisons were imprecise and uncertainty was substantial, particularly in sparse networks. CONCLUSIONS: GBL modalities may improve learning outcomes compared with TT, but relative effects appear domain-specific and the certainty of rankings is limited by sparse evidence and imprecision. Future trials should prioritise head-to-head comparisons, robust outcome measurement, and longer-term retention and transfer outcomes in both student and in-service populations.

Humans

Risk prediction models for blood transfusion in patients undergoing total hip and knee arthroplasty: a systematic review and meta-analysis.

OBJECTIVE: To systematically review and evaluate published risk prediction models for perioperative blood transfusion in patients undergoing total hip or knee arthroplasty (THA/TKA). METHODS: We systematically searched PubMed, Web of Science, the Cochrane Library, and Embase from inception to May 31, 2025. Two researchers independently screened the literature, extracted data, and assessed the risk of bias and applicability using the Prediction model Risk Of Bias Assessment Tool (PROBAST). The area under the receiver operating characteristic curve (AUC) values were pooled via a meta-analysis using Stata 18.0. RESULTS: d Fourteen studies containing 36 prediction models were included. The incidence of blood transfusion among THA/TKA patients ranged from 3.2% to 30.8%. Preoperative hemoglobin (Hb) level, tranexamic acid (TXA) use, operative duration, intraoperative blood loss, and age were the most frequently incorporated predictors. Model sensitivity ranged from 58% to 94.5%, and specificity ranged from 71.3% to 94%. Meta-analysis showed that the pooled AUC value of the 13 validated models was 0.87 (95% CI: 0.85-0.90), suggesting good discriminatory performance. All models were rated as having a high risk of bias. The applicability of four studies was rated as unclear. CONCLUSION: Although the included studies demonstrated promising discriminative ability of prediction models for blood transfusion in THA/TKA, all were assessed as having a high risk of bias using the PROBAST tool. Therefore, future research should prioritize the development of models with larger sample sizes, rigorous study designs, and multicenter external validation.

Humans

Machine learning-ready genomic biomarkers: ATF3 polymorphisms predict postoperative analgesic demand through AI-compatible phenotyping.

PURPOSE: To determine whether ATF3 polymorphisms can serve as genetic biomarkers for machine learning-based precision analgesia by establishing a genotype-phenotype association suitable for predictive modeling of postoperative opioid requirements. METHODS: In a prospective cohort of 167 adults undergoing abdominal surgery, ATF3 SNPs rs3122721 and rs3125293 were genotyped. A structured dataset architecture was developed to represent genetic profiles as input features for supervised learning models, enabling translational analysis of genotype&#x2011;dependent opioid consumption over 72&#xa0;h. RESULTS: Patients with homozygous genotypes of the ATF3 SNPs had significantly higher opioid requirements than non&#x2011;carriers, despite reporting similar subjective pain scores. This consistent genotype&#x2011;dependent pattern provided a clinically relevant phenotype suitable for integration into predictive algorithms. CONCLUSION: ATF3 genotyping offers a promising biomarker for computationally informed precision analgesia. By linking genomic variability to clinically meaningful outcomes within a structured clinical and genomic framework, this approach supports the future development of risk-stratified clinical decision-support systems to optimize postoperative pain management.Trial registration ChiCTR1900021991, registered 30 April 2019. SUPPLEMENTARY INFORMATION: The online version contains supplementary material available at https://doi.org/10.1007/s13755-026-00480-9.

ATF3

Data-centric, robust, and explainable multimodal deep learning for clinical decision support: A systematic review.

PURPOSE: Multimodal deep learning is increasingly proposed for clinical decision support (CDS) under a "data-centric" framing that prioritizes label quality, missing-modality robustness, distribution shift, calibration, and explainability. Prior reviews have examined multimodal medical AI, CDS, and data-centric methods separately, but none address their intersection. We mapped the modalities, fusion strategies, and data-centric and explainability techniques used in this recent literature, quantified how often each is implemented rather than merely mentioned, assessed deployment-relevant evidence (external validation, clinical-outcome measurement, equity), and formally appraised study-level risk of bias. METHODS: Following the PRISMA 2020 statement (PROSPERO CRD420261427815; registered retrospectively), we screened 150 records and included primary, clinical, multimodal studies that applied machine or deep learning to a decision-support task and reported at least one quantitative result. Two reviewers screened and extracted data with consensus adjudication. Each study was coded against pre-specified operational definitions, separating implemented or empirically evaluated techniques from those only mentioned. Study-level risk of bias was assessed with PROBAST + AI. Synthesis was narrative. RESULTS: Thirty-one studies met inclusion; 30 (97%) were published between 2024 and 2026, with a median of three modalities (range 2-6), most commonly structured EHR (71%) and imaging (39%). Data-centric techniques were frequently reported (74-84% across label-noise, distribution-shift, calibration, missing-modality and class-imbalance handling; equity 61%). However, external validation was reported in only 4/31 studies (13%), a clinical or provider outcome in 3/31 (10%), and no study reported routine deployment. Overall risk of bias was high in 27/31 studies (87%), driven by the analysis domain. CONCLUSION: Within this recent, self-selected slice of the field, technical robustness and explainability techniques are widely reported but rarely validated out-of-distribution or against clinical outcomes, and the underlying evidence is at high risk of bias. Progress requires external multi-site validation, clinical-outcome measurement, formal bias appraisal, and adherence to AI reporting standards (e.g., TRIPOD + AI) before deployment can be justified.

Deep Learning

A flipped classroom approach compared with low-interactive online learning for pediatric pain management knowledge and instructional motivation in nursing students: A randomized controlled study.

AIM: This study aimed to compare a flipped classroom approach with low-interactive online learning in terms of nursing students' questionnaire-assessed pediatric pain management knowledge and instructional motivation. BACKGROUND: Pain management in children is a critical and multidimensional nursing responsibility. However, limited curricular time and opportunities for applied learning may restrict nursing students' preparedness in this area. Structured and interactive instructional formats, such as the flipped classroom, may support knowledge acquisition and motivation in pediatric nursing education. METHODS: This study employed a parallel-group randomized controlled trial design with a 1:1 allocation ratio. Eighty-eight third-year prelicensure nursing students were randomized to either the flipped classroom group (n&#xa0;=&#xa0;44) or the low-interactive online learning group (n&#xa0;=&#xa0;44). Due to attrition (2 intervention, 2 control), analyses included 42 participants per group (n&#xa0;=&#xa0;84 in total). Data were collected between February and July 2022 using the Pediatric Pain Management Knowledge Scale for Nursing Students and the Instructional Materials Motivation Survey. This study was prospectively registered at ClinicalTrials.gov (Identifier: NCT07129044). RESULTS: At baseline, the groups were comparable in terms of knowledge and learning motivation. Following the intervention, the flipped classroom group demonstrated greater improvements in questionnaire-assessed pediatric pain management knowledge and instructional motivation than the low-interactive online learning group. Although scores declined from post-test to the three-month follow-up, they remained above baseline in the flipped classroom group. CONCLUSIONS: Within the context of this course, the flipped classroom approach was associated with greater improvement in questionnaire-assessed pediatric pain management knowledge and instructional motivation than low-interactive online learning. The findings should be interpreted as proximal educational outcomes rather than evidence of improved clinical competence or durable long-term effectiveness. Further studies using objective performance-based outcomes and longer follow-up periods are needed.

Humans

Artificial intelligence for dental caries detection: An umbrella review.

Artificial intelligence (AI) has been proposed as a tool to improve dental caries detection across imaging modalities; however, its clinical value remains uncertain. This umbrella review aimed to synthesize and critically appraise systematic reviews evaluating AI for caries detection and diagnosis. An umbrella review was conducted following PRIOR guidance (PROSPERO CRD420261340728). Searches were performed in MEDLINE, Embase, Scopus, Web of Science, and Google Scholar up to 15 March 2026. Methodological quality was assessed using AMSTAR 2, and overlap of primary studies was quantified using the corrected covered area (CCA). Seventeen systematic reviews were included, of which five reported diagnostic test accuracy meta-analyses using bivariate or HSROC models. Across these meta-analyses, pooled sensitivity ranged from 0.76 to 0.94 and specificity from 0.85 to 0.91. Most systems were based on deep learning models applied to bitewing radiographs and intraoral photographs. However, substantial heterogeneity was observed in imaging modalities, lesion thresholds, analytical tasks, and evaluation metrics. In addition, a high degree of overlap across reviews and recurrent methodological limitations, including reliance on retrospective datasets, limited external validation, and inconsistent reporting, substantially weaken the reliability of the evidence. Although AI models demonstrate high diagnostic performance under experimental conditions, current evidence does not support their use as stand-alone diagnostic tools. Their clinical applicability remains limited, and implementation should be restricted to decision-support contexts until robust prospective validation demonstrates meaningful impact on clinical decision-making and patient outcomes.

Dental Caries