Search PubMedSearch

SEARCH · Search PubMed

Results for “External validation”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

94 recordsLinked to original sources

Data-centric, robust, and explainable multimodal deep learning for clinical decision support: A systematic review.

PURPOSE: Multimodal deep learning is increasingly proposed for clinical decision support (CDS) under a "data-centric" framing that prioritizes label quality, missing-modality robustness, distribution shift, calibration, and explainability. Prior reviews have examined multimodal medical AI, CDS, and data-centric methods separately, but none address their intersection. We mapped the modalities, fusion strategies, and data-centric and explainability techniques used in this recent literature, quantified how often each is implemented rather than merely mentioned, assessed deployment-relevant evidence (external validation, clinical-outcome measurement, equity), and formally appraised study-level risk of bias. METHODS: Following the PRISMA 2020 statement (PROSPERO CRD420261427815; registered retrospectively), we screened 150 records and included primary, clinical, multimodal studies that applied machine or deep learning to a decision-support task and reported at least one quantitative result. Two reviewers screened and extracted data with consensus adjudication. Each study was coded against pre-specified operational definitions, separating implemented or empirically evaluated techniques from those only mentioned. Study-level risk of bias was assessed with PROBAST + AI. Synthesis was narrative. RESULTS: Thirty-one studies met inclusion; 30 (97%) were published between 2024 and 2026, with a median of three modalities (range 2-6), most commonly structured EHR (71%) and imaging (39%). Data-centric techniques were frequently reported (74-84% across label-noise, distribution-shift, calibration, missing-modality and class-imbalance handling; equity 61%). However, external validation was reported in only 4/31 studies (13%), a clinical or provider outcome in 3/31 (10%), and no study reported routine deployment. Overall risk of bias was high in 27/31 studies (87%), driven by the analysis domain. CONCLUSION: Within this recent, self-selected slice of the field, technical robustness and explainability techniques are widely reported but rarely validated out-of-distribution or against clinical outcomes, and the underlying evidence is at high risk of bias. Progress requires external multi-site validation, clinical-outcome measurement, formal bias appraisal, and adherence to AI reporting standards (e.g., TRIPOD + AI) before deployment can be justified.

Deep Learning

Diagnostic performance of machine learning models for malignant and non-malignant pleural effusion: Systematic review and meta-analysis.

BACKGROUND: Accurately distinguishing malignant pleural effusion (MPE) from non-malignant pleural effusion is clinically important, but the generalisability and methodological quality of machine-learning (ML) models remain uncertain. METHODS: We searched eight databases to 23 April 2026. Diagnostic performance was pooled using random-effects and Reitsma bivariate models, and study quality was assessed using PROBAST+AI. RESULTS: Forty-two studies were included; 17 contributed to the AUC meta-analysis and 14 to the bivariate analysis. The pooled AUC was 0.90 (95 % CI 0.85-0.94; 95 % prediction interval 0.62-0.98), with sensitivity of 0.80 (95 % CI 0.77-0.83) and specificity of 0.87 (95 % CI 0.79-0.92). Only nine studies reported external, temporal or independent validation. Externally validated studies had a lower pooled AUC than studies without external validation (0.83 vs 0.92), with lower specificity observed in the two externally validated studies contributing sensitivity and specificity data. All 42 development assessments had high overall quality concerns, and all 42 model evaluations were judged at high risk of bias. CONCLUSIONS: ML models showed good apparent accuracy for distinguishing MPE from non-MPE, but the evidence was limited by substantial heterogeneity, high risk of bias and scarce external validation. The pooled estimates reflect the average performance of different selected models rather than the expected accuracy of a single clinical test. ML models should be regarded as adjuncts to existing diagnostic pathways until they are confirmed by rigorous multicentre prospective external validation and clinical-impact studies.

Humans

Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.

BACKGROUND: Predictive modelling for cancer risk, treatment-related complications, and survival is central to precision oncology. Conventional logistic regression (LR) and Cox proportional hazards (CoxPH) regression remain widely used but are limited when modelling nonlinear interactions, high-dimensional imaging features, and multimodal clinical-metabolic predictors. Artificial intelligence (AI) and machine learning (ML) methods offer expanded capability through automated feature extraction, ensemble learning, and flexible survival modelling, but the evidence on when AI/ML adds value over conventional models across cancer sites and predictive tasks remains fragmented. OBJECTIVE: To systematically evaluate the methodological performance, validation strategies, and translational limitations of AI/ML models compared with conventional statistical models in published predictive-modelling studies for breast, colorectal, or pancreatic cancer. METHODS: PubMed, Scopus, and Web of Science were searched for studies published between January 2019 and March 2025. Two reviewers independently conducted title-and-abstract screening, full-text eligibility assessment, and PROBAST risk-of-bias assessment. Sixty-five studies (n = 907,567 participants) were narratively synthesised by cancer site, predictive task, model family, comparator, validation strategy, predictor modality, and calibration or explainability reporting. RESULTS: The 65 studies comprised breast cancer (n = 35), colorectal cancer (n = 21), and pancreatic cancer (n = 9). AI/ML superiority over LR and CoxPH was task- and data-dependent. CNN- and U-Net-based models predominated in imaging and body-composition tasks, tree-based ensembles consistently outperformed LR for tabular perioperative complication prediction, and CoxPH remained competitive, and in the largest pancreatic risk study, superior to XGBoost (C-index 0.802 vs 0.723) in well-structured datasets. PROBAST analysis-domain risk was moderate in 54 of 65 studies (83%), driven by limited external validation, sparse calibration reporting (11/65), and few decision-curve analyses (7/65). CONCLUSION: AI/ML adds the most methodological value in imaging-derived feature extraction and nonlinear perioperative prediction, while conventional regression remains preferable in large, structured datasets with linear predictors. Clinical translation requires standardised body-composition definitions, external validation, calibration assessment, decision-curve analysis, and explainability, in line with TRIPOD+AI and CLAIM standards.

Humans

Could the preoperative urethral curve be used to predict immediate urinary continence following Retzius-sparing robot-assisted radical prostatectomy? A retrospective multi-center study.

PURPOSE: Immediate urinary continence (UC) recovery following Retzius-sparing robot-assisted radical prostatectomy (RS-RARP) remains highly variable, highlighting the need for reliable preoperative prediction. We aimed to develop and validate models to identify patients likely to achieve immediate UC recovery following RS-RARP. MATERIALS AND METHODS: A total of 580 prostate cancer patients who underwent RS-RARP from four medical centers were assigned to a training set (n=348), an internal validation set (n=103) and an external validation set (n=129). Independent predictors were identified through univariate analysis and LASSO regression. A nomogram was constructed using multivariate logistic regression. Its performance was evaluated with receiver operating characteristic (ROC) curve, calibration curves, and decision curve analysis. RESULTS: Immediate UC recovery was observed in 84.5% (294/348) of patients in the training cohort, 80.6% (83/103) in the internal validation cohort, and 81.4% (105/129) in the external validation cohort, respectively. Multivariate analysis identified membranous urethral length (MUL) (OR=1.23, P=0.029) and urethral curvature (OR=2.84, P<0.001) as independent predictors, while prostate volume (PV) (OR=0.84, P <0.001) as a protective factor. The nomogram integrating MUL, PV, and urethral curvature demonstrated superior predictive accuracy, with an AUC of 0.87 (95% CI, 0.83-0.91) in the training cohort. The bootstrap-corrected calibration slope was 0.96, and the Brier score was 0.08.&#xa0;Calibration curves and decision curve analysis confirmed the predictive accuracy and clinical utility of the nomogram. CONCLUSIONS: Our study introduces a novel quantitative method for assessing urethral curvature. The mpMRI-based model, integrating urethral curvature and prostate spatial configuration, offers enhanced predictive accuracy for postoperative immediate UC recovery.

Humans

An individualized nomogram for predicting progression-free survival in systemic anaplastic large cell lymphoma: a multicenter, retrospective, and internally validated study.

OBJECTIVES: To develop an individualized nomogram for predicting disease progression risk in systemic anaplastic large cell lymphoma (sALCL). METHODS: Independent predictors of progression-free survival (PFS) were identified using Cox regression in a multicenter retrospective cohort of 109 sALCL patients (2010-2022). These were incorporated into a three-factor nomogram, evaluated via bootstrapped internal validation (1000 resamples), ROC analysis, C-index, decision curve analysis (DCA), and clinical impact curve (CIC). RESULTS: A total of 29 PFS events occurred during a median follow-up of 31 months. Multivariable modelling selected serum &#x3b2;2-microglobulin elevation, extranodal disease, and front-line chemotherapy choice (CHOP versus CHOPE or BV+CHP) as autonomous progression drivers. Upon internal bootstrap validation, the nomogram yielded strong prognostic accuracy, achieving AUCs of 0.81, 0.85 and 0.87 for 1-, 3- and 5-year progression-free survival, alongside a corrected C-index of 0.779 (95% CI: 0.699 - 0.861). Calibration plots showed close agreement between predicted and observed outcomes, while DCA confirmed superior net clinical benefit versus conventional IPI or Ann Arbor stratification across multiple decision thresholds. CONCLUSION: This first sALCL-specific nomogram integrates clinical and treatment variables to provide personalized PFS risk estimation. While internally validated, this exploratory, observation-based tool requires external validation and recalibration in prospective cohorts before clinical implementation.

Humans

A clinical study on the efficacy of rectal administration of Tongfu Qinghua decoction combined with external application of Ruyi Jinhuang powder in treating acute pancreatitis.

BACKGROUND: Acute pancreatitis (AP) is a common acute abdominal disease with high mortality in moderate and severe cases. Integrated Chinese and Western medicine therapy has promising clinical application prospects. OBJECTIVES: This study investigated the efficacy and safety of Tongfu Qinghua decoction enema combined with Ruyi Jinhuang powder external application for AP and its therapeutic effects across different age groups. METHODS: A total of 100 AP patients from October 2023 to August 2025 were randomly divided into observation and control groups (50 cases each). The control group received conventional Western medicine and the observation group received additional combined Chinese medicine therapy. Outcomes including hospital stay, symptom relief, inflammatory and pancreatic injury markers, clinical efficacy and adverse reactions were compared, with subgroup analysis of patients aged 18-40, 41-60 and 61-75 years. RESULTS: The observation group had significantly shorter hospital stay, faster symptom relief and gastrointestinal recovery (P<0.05). Post-treatment inflammatory and pancreatic markers improved significantly and the total effective rate was higher (P<0.05), with no significant difference in adverse reactions (P>0.05). Benefits were consistent across all age subgroups, with younger patients recovering faster and elderly patients still achieving significant improvement. CONCLUSION: This combined therapy is effective and safe for AP patients aged 18-75 years, significantly improving clinical outcomes and worthy of clinical promotion.

Humans

Applications of quantum AI in brain disorder diagnosis: A systematic review.

BACKGROUND AND OBJECTIVE: Brain disorder diagnosis and prediction remain challenging because neuroimaging, electrophysiological, behavioral, and multimodal data are high-dimensional, noisy, heterogeneous, and limited by small clinical cohorts. This systematic review synthesised applications of quantum artificial intelligence (QAI) for brain disorder diagnosis, prediction, detection, and monitoring. METHODS: Following PRISMA guidelines, studies published from 2016 to 13 January 2026 were retrieved from Scopus, Web of Science, and IEEE Xplore. After screening, 36 studies met the eligibility criteria and were qualitatively analysed according to disorder category, data modality, QAI method, implementation setting, validation strategy, and performance. RESULTS: At the broader disease-group level, neurodegenerative disorders were the most frequently investigated, followed by mental health and psychiatric disorders. At the individual level, Parkinson's disease and schizophrenia were the leading applications, followed by depression, anxiety, Alzheimer's disease, and stress-related tasks. MRI-based modalities were the most frequently used data source, followed by multimodal data and EEG. Methodologically, primary QAI approaches were dominated by quantum neural and QDL architectures, followed by quantum-inspired optimization or feature-selection methods and quantum-kernel/conventional QML classifiers. Qiskit/IBM Quantum and PennyLane were the most frequently reported quantum software frameworks. However, most studies relied on simulators, classical quantum-inspired implementations, or unclear implementation settings, with limited real-hardware evaluation. CONCLUSIONS: QAI shows emerging potential for brain disorder analysis, particularly through hybrid quantum-classical learning, quantum neural architectures, quantum-kernel methods, and quantum-inspired optimization. Nevertheless, current evidence remains preliminary and requires larger datasets, subject-level and external validation, fair classical benchmarking, noise-resilient circuits, real quantum hardware evaluation, explainability, and clinical validation.

Humans

Role of External Vibratory Lithoclast (Lithecbole) in Improving Lower Pole Stone Clearance After Extracorporeal Shock Wave Lithotripsy (ESWL).

<b>Background and Objective:</b> Extracorporeal shock wave lithotripsy (ESWL) is a widely used non-invasive treatment for renal stones. However, its effectiveness in clearing stones located in the lower pole of the kidney is often limited due to the anatomical challenges that impede the spontaneous passage of fragmented stones. This study evaluate the efficacy and safety of External Physical Vibration Lithecbole (EPVL) when used as an adjunctive therapy following ESWL in patients with lower pole renal stones, with a focus on improving stone clearance rates. <b>Materials and Methods:</b> This prospective interventional study was conducted at Al Yarmouk Teaching Hospital over two years and included 100 patients with 10-15 mm lower pole renal stones. Patients were randomized into two groups: The ESWL alone and ESWL plus EPVL. All patients received two ESWL sessions (3000 shocks/session). The treatment group additionally received EPVL therapy using a flank-applied vibration device. Follow-up was performed using ultrasound and KUB radiography. Statistical analysis was conducted using appropriate tests with significance set at p<0.05. <b>Results:</b> Baseline characteristics, including age, gender, weight, stone size and location, were statistically comparable between groups. The mean stone size was 12.4&#xb1;1.7 mm. A significantly higher stone-free rate was observed in the ESWL+EPVL group compared to the ESWL-only group (66% vs. 48%, p = 0.027). Although the residual stone size showed a numerical difference favoring EPVL, it was not statistically significant (p = 0.11). Complication rates were low, mild and similar in both groups (p = 0.3). Logistic regression analysis revealed that smaller stone size (p<0.001) and lower patient weight (p = 0.030) were significantly associated with successful stone clearance. Subgroup analysis further confirmed that patients with stones <11 mm or a weight <70 kg had the highest stone-free rates. <b>Conclusion:</b> In this initial evaluation, EPVL demonstrated a safe and effective adjunct to ESWL in enhancing stone clearance for lower-pole renal stones. Its application was associated with improved treatment outcomes without increasing complication rates. Further studies with extended EPVL sessions and longer follow-up are warranted.

Humans

Cross-tissue multi-omics integration highlights BPHL and mitochondrial targets in Alzheimer's disease.

BACKGROUND: Mitochondrial dysfunction is a hallmark of Alzheimer's disease (AD), yet specific molecular targets remain to be fully characterized. METHODS: A summary-data-based Mendelian randomization (SMR) framework integrated AD genome-wide association study (GWAS) statistics (39,918 cases) with blood DNA methylation quantitative trait loci (mQTL), gene expression (eQTL), and protein (pQTL) data for 1136 mitochondria-related genes. Associations were assessed using Bayesian colocalization and HEIDI testing. Tissue relevance was evaluated in four brain regions (hippocampus, amygdala, cortex, frontal cortex) using GTEx and external transcriptomic datasets. RESULTS: Screening identified eight candidates supported across blood mQTL and eQTL layers. Stepwise central nervous system (CNS) evaluation singled out biphenyl hydrolase-like (BPHL) as the consistent candidate. Higher genetically predicted BPHL expression was associated with reduced AD risk across the hippocampus (OR=0.920, 95% CI 0.873-0.970), amygdala (OR=0.925, 95%CI 0.880-0.973), cortex (OR=0.943, 95% CI 0.908-0.978), and frontal cortex (OR=0.938, 95%CI 0.901-0.976). These findings aligned with protein-protein interactions connecting BPHL to respiratory complexes and lower BPHL expression in independent AD brains. Functional enrichment converged on oxidative phosphorylation pathways. CONCLUSIONS: By integrating multi-omics data with tissue-specific validation, this study nominates BPHL as a consistent protective candidate in the brain. These findings provide genetic support for mitochondrial molecular perturbations in AD, offering insights for future validation.

Alzheimer Disease

Artificial intelligence for dental caries detection: An umbrella review.

Artificial intelligence (AI) has been proposed as a tool to improve dental caries detection across imaging modalities; however, its clinical value remains uncertain. This umbrella review aimed to synthesize and critically appraise systematic reviews evaluating AI for caries detection and diagnosis. An umbrella review was conducted following PRIOR guidance (PROSPERO CRD420261340728). Searches were performed in MEDLINE, Embase, Scopus, Web of Science, and Google Scholar up to 15 March 2026. Methodological quality was assessed using AMSTAR 2, and overlap of primary studies was quantified using the corrected covered area (CCA). Seventeen systematic reviews were included, of which five reported diagnostic test accuracy meta-analyses using bivariate or HSROC models. Across these meta-analyses, pooled sensitivity ranged from 0.76 to 0.94 and specificity from 0.85 to 0.91. Most systems were based on deep learning models applied to bitewing radiographs and intraoral photographs. However, substantial heterogeneity was observed in imaging modalities, lesion thresholds, analytical tasks, and evaluation metrics. In addition, a high degree of overlap across reviews and recurrent methodological limitations, including reliance on retrospective datasets, limited external validation, and inconsistent reporting, substantially weaken the reliability of the evidence. Although AI models demonstrate high diagnostic performance under experimental conditions, current evidence does not support their use as stand-alone diagnostic tools. Their clinical applicability remains limited, and implementation should be restricted to decision-support contexts until robust prospective validation demonstrates meaningful impact on clinical decision-making and patient outcomes.

Dental Caries

Integrated multi-omics profiling of amniotic fluid identifies predictive biomarkers for fetal growth restriction trajectories.

BACKGROUND: Fetal growth restriction (FGR) is a complex condition with highly heterogeneous clinical outcomes, making prenatal distinction between transient and persistent growth failure challenging. This study aims to identify amniotic fluid (AF) biomarkers capable of differentiating distinct FGR trajectories and characterizing persistent growth failure mechanisms. METHODS: Integrated proteomic and metabolomic profiling was performed on AF samples from transient FGR (n&#x2009;=&#x2009;11), persistent FGR (n&#x2009;=&#x2009;9), and healthy controls (n&#x2009;=&#x2009;13). Diagnostic and prognostic models were developed using multivariate analysis. Selected protein candidates were validated via ELISA in an independent cohort (n&#x2009;=&#x2009;69). RESULTS: Multi-omics analysis revealed distinct molecular signatures for FGR stratification. A two-protein diagnostic panel (PDGFA and phospho-STAT5A) achieved an AUC of 1.000 in the discovery stage and 0.780 in the external validation cohort. For prognostic assessment, a molecular signature including IREB2, HLA-C, and PLXNB2 accurately predicted persistent growth failure from transient recovery (AUC = 0.966). Cross-platform integration highlighted the mass spectrometry-derived WASHC2C as a central hub protein with a significant progressive increase across the control, transient, and persistent groups (p&#x2009;<&#x2009;0.001). CONCLUSIONS: This study establishes a multi-omics framework for prenatal FGR stratification. Our findings identify distinct molecular&#xa0;signatures reflecting&#xa0;the intrauterine environment and provide high-performance molecular tools for predicting divergent fetal growth trajectories to guide personalized clinical decision-making.

Humans

Risk prediction models for blood transfusion in patients undergoing total hip and knee arthroplasty: a systematic review and meta-analysis.

OBJECTIVE: To systematically review and evaluate published risk prediction models for perioperative blood transfusion in patients undergoing total hip or knee arthroplasty (THA/TKA). METHODS: We systematically searched PubMed, Web of Science, the Cochrane Library, and Embase from inception to May 31, 2025. Two researchers independently screened the literature, extracted data, and assessed the risk of bias and applicability using the Prediction model Risk Of Bias Assessment Tool (PROBAST). The area under the receiver operating characteristic curve (AUC) values were pooled via a meta-analysis using Stata 18.0. RESULTS: d Fourteen studies containing 36 prediction models were included. The incidence of blood transfusion among THA/TKA patients ranged from 3.2% to 30.8%. Preoperative hemoglobin (Hb) level, tranexamic acid (TXA) use, operative duration, intraoperative blood loss, and age were the most frequently incorporated predictors. Model sensitivity ranged from 58% to 94.5%, and specificity ranged from 71.3% to 94%. Meta-analysis showed that the pooled AUC value of the 13 validated models was 0.87 (95% CI: 0.85-0.90), suggesting good discriminatory performance. All models were rated as having a high risk of bias. The applicability of four studies was rated as unclear. CONCLUSION: Although the included studies demonstrated promising discriminative ability of prediction models for blood transfusion in THA/TKA, all were assessed as having a high risk of bias using the PROBAST tool. Therefore, future research should prioritize the development of models with larger sample sizes, rigorous study designs, and multicenter external validation.

Humans

PGR expression as a pharmacogenomic companion biomarker to GENE70-derived genomic risk in ER-positive/HER2-negative breast cancer.

BACKGROUND: The biology of the estrogen receptor-positive (ER+) and human epidermal growth factor receptor 2-negative (HER2-) breast cancers is heterogeneous even when they are categorized by their risk via genomics. Transcriptomic PGR expression reflects endocrine pathway activity and may provide complementary biological information within established GENE70-derived genomic-risk categories. Whether this molecular marker improves the biological interpretation of genomic-risk stratification beyond conventional clinicopathological assessment remains uncertain. OBJECTIVES: The aim of this study was to determine whether transcriptomic PGR expression provides complementary biological and prognostic information within reconstructed GENE70-derived genomic-risk categories and refines the characterization of endocrine-related tumour biology in ER-positive/HER2-negative breast cancer. METHODS: This study analysed publicly available transcriptomic and clinical data from three cohorts: METABRIC (discovery cohort), GSE96058/SCAN-B cohort (validation cohort) and TCGA-BRCA cohort (molecular validation cohort). The GENE70-derived genomic-risk score was reconstructed for each cohort using matched genes. Cox regression, Kaplan-Meier analysis and subgroup comparisons were used to assess relationships between PGR expression, clinicopathologic variables, molecular features and survival outcomes. RESULTS: Across the three independent cohorts, low transcriptomic PGR expression was consistently associated with higher GENE70-derived genomic risk, increased MKI67 expression, reduced ESR1 expression and enrichment of the Luminal B subtype. Survival findings differed between cohorts. In the discovery METABRIC cohort, transcriptomic PGR expression showed heterogeneous associations with survival, particularly within GENE70-derived high-risk subgroups, whereas the external GSE96058/SCAN-B validation cohort demonstrated consistent associations between low PGR expression and poorer overall survival in both the overall ER-positive/HER2-negative population and GENE70-derived high-risk subgroups. CONCLUSION: These findings suggest that transcriptomic PGR provides complementary biological and prognostic information within GENE70-derived genomic-risk categories. However, because treatment response was not evaluated in the present study, the findings should not be interpreted as evidence of predictive or pharmacogenomic utility and prospective studies incorporating treatment-response analyses are required before such applications can be established.

Humans

Impact of estimated total blood volume on NT-proBNP response to angiotensin receptor-neprilysin inhibition in acute heart failure: Insights from the PREMIER study.

BACKGROUND: Sacubitril/valsartan (Sac/Val) reduces N-terminal pro-B-type natriuretic peptide (NT-proBNP) levels in acute heart failure (AHF), particularly in patients with reduced ejection fraction. However, whether estimated total blood volume (TBV), calculated using anthropometric equations, is associated with heterogeneity in biomarker response remains uncertain. METHODS: This post hoc exploratory sub-analysis of the PREMIER randomized trial evaluated whether baseline estimated TBV was associated with heterogeneity in NT-proBNP reduction after Sac/Val compared with angiotensin-converting enzyme inhibitor/angiotensin receptor blocker (ACEI/ARB) therapy. Estimated TBV was calculated using validated anthropometric equations and dichotomized at the median (4.05 L). Patients were further stratified by left ventricular ejection fraction (LVEF <40% vs &#x2265;40%). The primary endpoint was the proportional change in NT-proBNP from baseline to Week 8. RESULTS: Among 376 patients, 372 with baseline estimated TBV data were analyzed. In the high TBV group, Sac/Val was associated with greater NT-proBNP reduction than ACEI/ARB (-56% vs -32%; ratio of change, 0.67; 95% confidence interval, 0.53-0.84; P = .001), whereas no significant difference was observed in the low TBV group (P for heterogeneity = 0.063). In patients with LVEF <40%, Sac/Val was associated with greater NT-proBNP reduction in both TBV groups. In patients with LVEF &#x2265;40%, Sac/Val was associated with greater NT-proBNP reduction in the high TBV group, whereas the point estimate in the low TBV group numerically favored ACEI/ARB. CONCLUSIONS: In this exploratory post hoc analysis, higher estimated TBV was associated with greater NT-proBNP reduction after Sac/Val, particularly among patients with LVEF &#x2265;40%. These findings are hypothesis-generating and require external validation. TRIAL REGISTRATION: ClinicalTrials.gov, NCT05164653; Japan Registry of Clinical Trials, jRCTs021210046.

Humans

Association of time-averaged systemic immune-inflammation indices with in-hospital mortality after intracerebral hemorrhage: a retrospective study.

BACKGROUND: Systemic inflammation plays a central role in secondary brain injury following intracerebral hemorrhage (ICH). Although inflammatory indices such as the neutrophil-to-lymphocyte ratio (NLR), systemic immune-inflammation index (SII), and systemic inflammation response index (SIRI) are linked to poor outcomes, their associations with mortality are commonly assumed to be linear, potentially overlooking nonlinear patterns where mortality risk rises steeply at higher levels. METHODS: We conducted a retrospective study using the MIMIC-IV database, including 440 patients with non-traumatic ICH who were alive and remained in the ICU for at least 72&#xa0;h after admission. Mean NLR, SII, and SIRI were calculated from measurements obtained during this period. Multivariable logistic regression and restricted cubic spline (RCS) analyses were applied to assess their independent and nonlinear associations with in-hospital mortality. Model discrimination and calibration were internally validated using 1,000 bootstrap resamples. RESULTS: The in-hospital mortality rate was 26.1%. After multivariable adjustment, NLR and SIRI remained independently associated with mortality. Patients in the highest SIRI quartile had the highest risk of death (aOR&#xa0;=&#xa0;5.12; 95% CI: 2.57-12.24; p&#xa0;<&#xa0;0.001). RCS analysis revealed a significant nonlinear association between SIRI and mortality (p-nonlinearity&#xa0;<&#xa0;0.05), showing a steep risk increase at higher SIRI levels. Adding SIRI to the base model provided a modest improvement in discrimination (AUC 0.762 to 0.785, p&#xa0;=&#xa0;0.045) and significantly improved risk reclassification (cNRI&#xa0;=&#xa0;0.4778, p&#xa0;<&#xa0;0.001; IDI&#xa0;=&#xa0;0.0240, p&#xa0;=&#xa0;0.0151). CONCLUSIONS: Among patients with ICH who met the 72-hour eligibility criterion, higher 72-hour average SIRI was independently associated with in-hospital mortality. As a time-averaged measure, SIRI should be interpreted as a dynamic marker integrating the initial inflammatory state and the early clinical course rather than as a purely baseline prognostic factor. Although adding SIRI to the base model modestly improved discrimination and risk reclassification, it should be considered a candidate prognostic marker requiring external validation before clinical application.

Humans

ALID score for treatment-effect heterogeneity of adjunctive low-voltage area ablation in persistent atrial fibrillation: A post hoc analysis of SUPPRESS-AF.

BACKGROUND: In persistent atrial fibrillation (AF), the incremental benefit of adjunctive low-voltage area (LVA) ablation beyond pulmonary vein isolation (PVI) remains inconsistent. OBJECTIVE: To examine whether a simple clinical score characterizes treatment-effect heterogeneity of adjunctive LVA ablation among patients with mapped LVA&#xa0;>&#xa0;5&#xa0;cm2 and to perform an exploratory supportive analysis in an independent randomized cohort. METHODS: In this post-hoc analysis of SUPPRESS-AF, which included patients with persistent AF and mapped LVA&#xa0;>&#xa0;5&#xa0;cm2 after PVI, four variables-age&#xa0;&#x2265;&#xa0;75&#xa0;years, left atrial diameter&#xa0;>&#xa0;44&#xa0;mm, estimated glomerular filtration rate&#xa0;<&#xa0;60&#xa0;mL/min/1.73&#xa0;m2, and absence of diabetes-were combined into the ALID score (0-4). Patients were stratified into low (0-1), intermediate (2), and high (3-4) score groups. Because EARNEST-PVI did not use LVA-guided ablation or select patients based on mapped LVA, it was analyzed as an exploratory supportive cohort rather than as an external validation cohort. RESULTS: In SUPPRESS-AF (n&#xa0;=&#xa0;336), a significant treatment-by-score interaction was observed (P&#xa0;<&#xa0;0.001). Adjunctive LVA ablation was associated with increased recurrence in the low-score stratum (HR 3.92; 95% CI 1.50-10.20) and reduced recurrence in the high-score stratum (HR 0.48; 95% CI 0.26-0.86). In EARNEST-PVI (n&#xa0;=&#xa0;494), a qualitatively similar interaction pattern was observed for additional ablation beyond PVI (interaction P&#xa0;=&#xa0;0.029), although the ablation strategy differed from LVA-guided ablation. CONCLUSIONS: Among patients with persistent AF and mapped LVA >5&#xa0;cm2, the ALID score identified heterogeneity in response to adjunctive LVA ablation. These hypothesis-generating findings require prospective validation before clinical implementation.

Humans

Quo vadis, BGA? A collaborative EDNAP exercise on the challenges and progress in forensic biogeographical ancestry inference.

There is a broad consensus that forensic tests for the prediction of externally visible characteristics (EVC) and analysis of biogeographic ancestry (BGA) of an individual are technically reliable. However, interpretation of the results and population-specific genotype distribution patterns remains challenging. EVC and BGA analyses provide valuable information for population genetics studies and as investigative leads for criminal cases, as well as for historical and contemporary identification tests. However, inaccurate or incorrect predictions, for example, from subjective bias in the interpretations made, have the potential to misdirect police investigations. The legal situation regarding EVC and BGA testing varies by country: ranging from countries where it is explicitly prohibited, to those without specific regulations on biogeographic ancestry prediction, and others that have already enacted laws governing its use. The reluctance to utilize these analyses is not only due to legal restrictions and data protection concerns, but also to initial limited sets of sufficiently comprehensive forensic DNA assays. Forensic BGA marker panels typically contain up to &#x223c;300 SNPs. This relatively small number of genetic markers, along with limited reference population data, complicates the interpretation of results from donors of unknown origin. This paper presents the results of a collaborative EDNAP study, which, for the first time, evaluated the approach to reporting EVC and BGA data between international laboratories. For the study, DNA from nine individuals with self-reported ancestry was collected and analysed using various forensic panels differing in the number and composition of ancestry-informative markers genotyped, comprising: the Precision ID mtDNA Whole Genome Panel, the VISAGE Basic Tool and the VISAGE Enhanced Tool for Appearance and Ancestry Prediction, and the Ion AmpliSeq&#x2122; PhenoTrivium Panel. To ensure full data protection, all SNP genotypes and uniparental marker haplotypes obtained were not shared with third parties. Instead, the genetic data were analysed using a range of commonly used population analysis software packages. These analysis outcomes were then distributed to twelve European forensic laboratories (both academic and law enforcement institutions), who were asked to prepare reports based on their interpretation of the phenotypes and ancestry they inferred from the analysis data. A questionnaire sent alongside the genetic information, aimed to evaluate which difficulties were encountered by the participants in processing the BGA analysis data they were given.

Humans

Perinatal depression, maternal thyroid status and fetus/infant health and development: A systematic review.

BACKGROUND: Thyroid hormones are known to influence both maternal depression and child developmental outcomes, while maternal depression independently affects child outcomes. The potential interaction between thyroid dysfunction and depression in shaping child development remains insufficiently explored. The present study addresses such interplay. METHODS: Following PRISMA 2020 and JBI guidelines, three databases were searched through December 2025 for primary studies on maternal thyroid status, perinatal depression, and child development. Risk of bias (RoB) was assessed using validated tools. Due to clinical and methodological heterogeneity, data were synthesized narratively following SWiM guidelines. RESULTS: Eleven studies were included. Beyond independent risks for preterm birth and behavioral problems, limited evidence supports a synergistic model, while most studies likely reflect the simple co-occurrence of risks. Maternal thyroid peroxidase antibodies (TPO-Ab) were associated with child externalizing problems exclusively in the presence of clinical depression. High depressive symptoms also attenuated the cognitive benefits of prenatal iodine supplementation. Thyroid status appears to function as a risk moderator rather than a mediator. However, 50% of observational studies presented high RoB, primarily due to participant attrition. CONCLUSION: Findings are still scarce to support a synergistic risk model where specific maternal thyroid parameters (i.e. thyroid autoimmunity and iodine status) may moderate the impact of depressive symptoms on child development. Despite the high RoB in half of the studies, results highlight the need for integrated screening protocols. Simultaneously assessing mental health and thyroid status may optimize risk stratification for high-risk mother-infant dyads.

Female