Search PubMedSearch

SEARCH · Search PubMed

Results for “Validation Studies as Topic”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

2,791 records · Page 14Linked to original sources

Characteristics of p53 and Smad4 immunohistochemistry in pancreatic ductal adenocarcinoma and validation by next-generation sequencing.

BACKGROUND: Mutations in four major driver genes -KRAS, CDKN2A, TP53, and SMAD4- are central to the pathogenesis of pancreatic ductal adenocarcinoma (PDAC) and critically inform diagnosis, therapeutic decision-making, and prognostic assessment. Although next-generation sequencing (NGS) is widely regarded as the gold standard for detecting these mutations, its clinical application is often limited by suboptimal analytical efficiency and substantial economic cost. Among these genes, immunohistochemical (IHC) staining for the proteins encoded by TP53 and SMAD4 has been extensively adopted in routine pathology practice. However, standardized IHC pattern classification schemes and rigorous validation of their predictive accuracy for underlying genomic alterations remain lacking in PDAC. METHODS: We retrospectively enrolled 63 PDAC patients and systematically characterized the typical IHC expression patterns of p53 and Smad4. Targeted NGS was subsequently performed on all available tumor specimens, and the resulting mutational profiles were correlated with corresponding IHC findings. Diagnostic performance including sensitivity, specificity and accuracy of p53 IHC for predicting TP53 mutations and of Smad4 IHC for predicting SMAD4 mutations was rigorously evaluated. RESULTS: Among the four canonical driver genes, co-occurring double- or triple-gene mutations were prevalent; within TP53 and SMAD4, missense mutations constituted the most frequent variant type. Using NGS as the reference standard, we validated the diagnostic utility of a three-tiered p53 IHC classification system, particularly in fine-needle biopsy (FNB) specimens. Furthermore, we proposed a novel, refined Smad4 IHC pattern classification that incorporates an "intermediate" category, thereby expanding upon conventional binary interpretation. This new scheme achieved markedly improved mutation prediction accuracy (0.76) compared with traditional approaches (0.57). CONCLUSION: Our study highlights the complementary diagnostic value of p53 and Smad4 IHC relative to molecular testing in PDAC, especially when tissue is limited, as commonly encountered in FNB specimens. The newly established Smad4 IHC classification system, which integrates an intermediate expression category into the conventional two-tier framework, demonstrates superior clinical utility and enhances predictive accuracy for SMAD4 genomic alterations.

Humans

Genome-wide identification and functional validation of asparagine synthetase genes (NtASNs) in Nicotiana tabacum.

Asparagine (Asn) is pivotal for plant nitrogen (N) metabolism and plays indispensable roles in plant growth, development, and stress tolerance. However, the systematic characteristics and core functions of asparagine synthetase genes (NtASNs) in tobacco remain unclear. Through a comprehensive genome-wide investigation, nine members of the NtASN gene family were identified. Subsequent CRISPR/Cas9-mediated knockout and overexpression assays of these NtASN genes revealed that NtASN1e, NtASN2a, and NtASN2b are the core genes responsible for Asn biosynthesis in tobacco. Their knockout reduced asparagine synthetase activity and Asn content, delayed seed germination by 2-3 days, and displayed elevated oxidative injury when exposed to salinity conditions. In contrast, overexpression of these genes elevated Asn accumulation. Subcellular localization analysis indicated that NtASN1e was localized to both the cytoplasm and chloroplasts, whereas NtASN2a exhibited dual localization in the cytoplasm and endoplasmic reticulum, and NtASN2b was mainly localized in the cytoplasm. This study systematically clarifies the evolutionary characteristics and core functions of the NtASN gene family and provides candidate genes for optimizing nitrogen metabolism and improving salt-stress adaptation in tobacco. These findings hold important practical significance for molecular breeding and product quality improvement in industrial crops.

Nicotiana

Olive leaf protein hydrolysates yield gastro-resistant peptides with antioxidant and anti-inflammatory potential: peptidomics, in vitro validation and molecular docking analyses.

Olive (Olea europaea L.) leaves are an abundant olive-oil by-product and a promising feedstock for sustainable valorisation. An olive leaf protein isolate (OLPI) from olive-leaf powder (OLP) was enzymatically hydrolysed to yield seven hydrolysates (OLPHs). All showed notable antioxidant activity as whole hydrolysate matrices (EC₅₀ = 0.11-0.28 mg mL-1); likely reflecting the combined contribution of released peptides and co-extracted phenolic compounds; the 15-min Alcalase product (OLPH15A) showed high activity with the shortest processing time. Its INFOGEST digest (dOLPH15A) attenuated LPS-induced inflammation in Caco-2 cells, down-regulating pro-inflammatory and up-regulating anti-inflammatory genes. Peptidomics identified 7037 peptides in OLPH15A and 534 in dOLPH15A, from which twenty gastro-resistant sequences were prioritised for in silico analysis. Multi-tool prediction and docking highlighted four peptides, GAAGGIGQPL, QSAYPGTGPL, GGGAGGGDGGIL and LDAQFPGVN, with favourable predicted affinity for the TLR4/MD2 complex, suggesting that they may contribute to the observed immunomodulatory response. These findings position olive leaves as a viable source of protein hydrolysate-based ingredients with antioxidant and anti-inflammatory potential, advancing the valorisation of olive-oil by-products.

Olea

Phytolacca acinosa Roxb. induces intestinal toxicity through the histamine-MLCK-tight junction axis: Integrated evidence from proteomics, metabolomics, intestinal organoids and epithelial barrier validation.

Phytolacca acinosa Roxb. (PR) is a saponin-rich medicinal plant associated with gastrointestinal toxicity, but the mechanisms underlying PR-induced intestinal barrier injury remain unclear. In this study, raw PR extract was analytically characterized by UPLC-ZenoTOF-MS/MS, confirming triterpenoid saponins as the predominant constituents. C57BL/6 J mice were orally exposed to characterized PR extract (1.20 or 12.0 g/kg for 5 h), and Caco-2 cells and mouse intestinal organoids were used to assess epithelial toxicity and barrier disruption. Histopathology, ELISA, FITC-dextran permeability assays, immunofluorescence, CCK-8, LDH release, western blotting, DIA-based proteomics and untargeted metabolomics were integrated to define toxicological mechanisms. PR induced dose-dependent intestinal inflammation and barrier dysfunction, with the ileum as the most sensitive target. PR increased serum DAO and D-lactate and intestinal TNF-α and IL-1β, disrupted organoid morphology, enhanced epithelial permeability, and reduced ZO-1 expression. Proteomics revealed changes in inflammatory, lipid-metabolic, cytoskeletal and tight-junction pathways, including upregulation of MLCK3 and phospholipase-related proteins and downregulation of ZO-1 and ZO-2. Metabolomics identified histidine metabolism disturbance and histamine accumulation. Integrated multi-omics and pharmacological validation indicated that histamine activated the PLC/IP₃/Ca²⁺/CaM/MLCK cascade, promoting MLC phosphorylation, tight-junction disassembly and epithelial leakiness. MLCK inhibition partially restored ZO-1/ZO-2 expression and attenuated PR-induced epithelial injury. These findings identify the histamine-MLCK-tight junction axis as a key mechanism of PR-induced intestinal toxicity and support hazard identification of saponin-rich PR exposure.

Animals

Comprehensive analysis of mRNA-microRNA-lncRNA expression profiles in post-traumatic elbow heterotopic ossification using RNA sequencing and experimental validation.

BACKGROUND: This study aimed to profile the molecular signatures of post-traumatic elbow heterotopic ossification (HO) to identify key regulators and potential therapeutic targets. METHODS: Total RNA from post-traumatic elbow HO tissues (n=4) and normal bone tissues (n=6) was subjected to high-throughput sequencing to identify differentially expressed mRNAs (DEGs), microRNAs (DEMs), and lncRNAs (DELs). Bioinformatics analyses included Gene Ontology (GO), Kyoto Encyclopedia of Genes and Genomes (KEGG) pathway enrichment, protein-protein interaction network construction, and transcription factor (TF)-microRNA-mRNA network analysis. The expression trends of four most upregulated and four most downregulated DEGs were validated by real-time quantitative reverse transcription polymerase chain reaction (qRT-PCR). RESULTS: We identified 2,138 DEGs, 40 DEMs, and 905 DELs. DEGs were significantly enriched in biological process "bone mineralization," cellular component "plasma membrane," molecular function "integrin binding," and pathways including PI3K-Akt, NF-κB, JAK-STAT, and TNF signaling pathways. Hub genes with high connectivity included MMP9, IL6, MMP3, CTSK, and BGLAP. Integrated network analysis highlighted the transcription factor JUN and key microRNAs (hsa-miR-124-3p, hsa-miR-548c-3p, and hsa-miR-135b). The qRT-PCR results confirmed the expression trends of selected DEGs. CONCLUSIONS: This study, for the first time, profiled the differentially expressed mRNAs, microRNAs, and lncRNAs in post-traumatic elbow HO using high-throughput RNA sequencing. These findings provide valuable insights into the molecular mechanisms of HO following elbow trauma. The identified hub genes (MMP9, IL6, MMP3, CTSK, and BGLAP), key TF (JUN), and key microRNAs (hsa-miR-124-3p, hsa-miR-548c-3p, and hsa-miR-135b) may serve as potential therapeutic targets for preventing and treating post-traumatic elbow HO.

Humans

Dynamic evolution of chaperone-mediated autophagy is associated with tumor microenvironment remodeling and prognostic stratification in lung adenocarcinoma: insights from single-cell transcriptomics, ensemble machine learning, and experimental validation.

BACKGROUND: Lung adenocarcinoma (LUAD) shows prognostic heterogeneity, and tumor-node-metastasis (TNM) staging is limited for individualized management. Chaperone-mediated autophagy (CMA) maintains proteostasis, but its role during adenocarcinoma in situ (AIS)-minimally invasive adenocarcinoma (MIA)-invasive adenocarcinoma (IAC) progression remains unclear. METHODS: Single-cell RNA sequencing (scRNA-seq) data from GSE189357 and bulk transcriptomes from The Cancer Genome Atlas (TCGA)-LUAD and Gene Expression Omnibus (GEO) cohorts were integrated. CMA activity, cell-cell communication, weighted gene co-expression network analysis (WGCNA), tumor-normal differential expression, machine-learning survival modeling, tumor microenvironment (TME) features, drug sensitivity, and EPC1 function were analyzed. RESULTS: CMA-high tumor epithelial cells increased from AIS (58.1%) to MIA (65.7%) but declined in IAC (44.4%; p < 0.001). CMA-low cells preferentially received fibroblast-derived extracellular matrix cues. A CMA-negatively correlated module identified 69 core genes. Random survival forest (RSF) performed best among 117 machine-learning combinations (mean concordance index > 0.873). High-risk patients had worse survival across cohorts, and the risk score was independently associated with overall survival (hazard ratio = 16.013, 95% confidence interval: 9.579-26.768, p < 0.001). High-risk tumors showed proliferative activation and M0 macrophage enrichment, whereas low-risk tumors showed stronger immune-related signaling. EPC1 overexpression suppressed malignant phenotypes in A549 cells. CONCLUSION: CMA dynamics are associated with stromal and immune remodeling during LUAD progression. A CMA-based model provides robust prognostic stratification and may offer a basis for future TME-guided studies.

Chaperone-mediated autophagy

Stretched penile length in boys with hypospadias: Population-based analysis using validated nomogram.

BACKGROUND: Hypospadias affects 1 in 200-300 male births. Parents are often concerned about penile adequacy beyond the urethral defect itself, yet few studies have systematically compared stretched penile length (SPL) in hypospadias against population-based reference standards. OBJECTIVE: To evaluate SPL distribution patterns in boys with Types I and II hypospadias and compare them with established normative data. METHODS: The authors studied 876 consecutive boys aged 1-14 years with unoperated Types I (distal) and II (mid-shaft) hypospadias. Two observers independently measured SPL using the validated SPLINT technique. The SPL measurements were compared against age-matched normative data from 1276 Indian children. Exact binomial probability tests were used for percentile distributions, chi-square tests for subtype comparisons and t-tests for mean deviations. RESULTS: The cohort included 479 Type I and 397 Type II cases. SPL distribution showed a marked leftward shift: 71% fell below the 50th percentile (expected 50%, p < 0.001) and 41.5% below the 25th percentile. Lower percentiles were overrepresented, 20.7% were below the 10th percentile and 20.8% in the 10th-25th range. Upper percentiles were depleted: only 7.4% in the 75th-90th range and 1.7% above the 90th percentile (all p < 0.001). Mean SPL was reduced by 6.8% (95% CI: -8.18 to -5.42%) in Type I and 7.5% (95% CI: -9.05 to -5.92%) in Type II. The two subtypes showed no significant distributional difference (&#x3c7;2 = 6.22, p = 0.18), suggesting that meatal position does not predict SPL reduction. CONCLUSIONS: Boys with distal and mid-shaft hypospadias show clinically meaningful SPL reduction that follows a continuous distribution rather than an all-or-none pattern. SPL reduction appears independent of meatal position. These findings support routine SPL assessment using population-specific references and can guide preoperative counselling.

Humans

Validation of the newly introduced Deauville score 5a for patients treated for advanced-stage classic Hodgkin lymphoma.

The Lugano Imaging Committee recently refined the Deauville score (DS), subdividing DS5 into DS5a (>2&#xd7; liver uptake without new lesions) and DS5b (new lesions). We investigated whether this improves prognostic discrimination at interim positron emission tomography (PET) after 2 cycles (PET-2) in patients with advanced-stage classical Hodgkin lymphoma (AS-cHL) treated in recent German Hodgkin Study Group randomized phase 3 trials. The primary analysis cohort was HD18 postamendment standard arms (uniform treatment with 6 cycles of escalated doses of bleomycin, etoposide, doxorubicin, cyclophosphamide, vincristine, procarbazine, and prednisone [eBEACOPP]); sensitivity cohorts were HD18 intention-to-treat and HD21 eBEACOPP and brentuximab vedotin, etoposide, cyclophosphamide, doxorubicin, dacarbazine, and dexamethasone arms. Progression-free survival (PFS) was analyzed by landmark Cox models starting at PET-2. DS5a was infrequent (4%-6% across cohorts; 39/639, 67/1745, 33/568, and 29/560). In the primary cohort, DS5 was associated with inferior PFS vs DS1 to DS3 (hazard ratio [HR], 3.00; 95% confidence interval [CI], 1.25-7.23) and vs DS1 to DS4 (HR, 2.35; 95% CI, 1.01-5.50). Across sensitivity cohorts, DS5a remained adverse compared with DS1 to DS4 (HR range, 2.57-5.47), whereas DS4 according to the new definition did not consistently separate from DS1 to DS3, which is likely a result of PET-adapted treatment. Overall survival trends were concordant, but interpretation is limited by few events. To our knowledge, this is the first prognostic validation of the refined DS in prospectively randomized trial populations. The newly introduced DS5a isolates a small high-risk AS-cHL, which further supports risk assessment and adaptation using quantitative biomarkers from PET. The HD18 and HD21 trials were registered at www.clinicaltrials.gov as NCT00515554 and NCT02661503, respectively.

Humans

Could the preoperative urethral curve be used to predict immediate urinary continence following Retzius-sparing robot-assisted radical prostatectomy? A retrospective multi-center study.

PURPOSE: Immediate urinary continence (UC) recovery following Retzius-sparing robot-assisted radical prostatectomy (RS-RARP) remains highly variable, highlighting the need for reliable preoperative prediction. We aimed to develop and validate models to identify patients likely to achieve immediate UC recovery following RS-RARP. MATERIALS AND METHODS: A total of 580 prostate cancer patients who underwent RS-RARP from four medical centers were assigned to a training set (n=348), an internal validation set (n=103) and an external validation set (n=129). Independent predictors were identified through univariate analysis and LASSO regression. A nomogram was constructed using multivariate logistic regression. Its performance was evaluated with receiver operating characteristic (ROC) curve, calibration curves, and decision curve analysis. RESULTS: Immediate UC recovery was observed in 84.5% (294/348) of patients in the training cohort, 80.6% (83/103) in the internal validation cohort, and 81.4% (105/129) in the external validation cohort, respectively. Multivariate analysis identified membranous urethral length (MUL) (OR=1.23, P=0.029) and urethral curvature (OR=2.84, P<0.001) as independent predictors, while prostate volume (PV) (OR=0.84, P <0.001) as a protective factor. The nomogram integrating MUL, PV, and urethral curvature demonstrated superior predictive accuracy, with an AUC of 0.87 (95% CI, 0.83-0.91) in the training cohort. The bootstrap-corrected calibration slope was 0.96, and the Brier score was 0.08.&#xa0;Calibration curves and decision curve analysis confirmed the predictive accuracy and clinical utility of the nomogram. CONCLUSIONS: Our study introduces a novel quantitative method for assessing urethral curvature. The mpMRI-based model, integrating urethral curvature and prostate spatial configuration, offers enhanced predictive accuracy for postoperative immediate UC recovery.

Humans

Effects of time-restricted eating on markers of glucose metabolism and regulation in individuals with prediabetes or type 2 diabetes: a systematic review and meta-analysis of randomised controlled trials.

AIMS/HYPOTHESIS: This systematic review and meta-analysis aimed to investigate the effects of time-restricted eating (TRE) on glucose metabolism and regulation in individuals with prediabetes (fasting blood glucose of 5.6-6.9 mmol/l or HbA1c of 39-47 mmol/mol [5.7-6.4%]) or type 2 diabetes (fasting blood glucose &#x2265;7 mmol/l or HbA1c &#x2265;48 mmol/mol [6.5%]). METHODS: A literature search was performed in MEDLINE, Embase and CENTRAL from inception to 5 August 2025. Moreover, forward and backward citation searches were performed. Eligible studies were RCTs in adults with prediabetes or type 2 diabetes, lasting &#x2265;2 weeks, reporting markers of glucose metabolism and regulation, comparing TRE (&#x2264;12 h eating window) with a non-time-restricted control diet. Studies involving pregnancy, other fasting regimens, or non-peer-reviewed publications were excluded. Data were pooled as weighted mean differences with 95% CIs using random-effects generic inverse variance models in Cochrane Review Manager Web, and results are presented as forest plots. The certainty of evidence was defined using Grading of Recommendations, Assessment, Development and Evaluations methodology, and risk of bias was estimated by using the Revised Cochrane risk-of-bias tool for randomised trials (RoB 2). RESULTS: Out of 2043 records identified through the database search, as well as 1249 from forward and backward citation searches, ten RCTs including 599 participants were included. The mean length of the studies was 4 months, and the eating windows ranged from 4 to 10 h per day. The pooled meta-analysis showed no overall effect of TRE on HbA1c (-3.33 mmol/mol; 95% CI -6.87, 0.20 (-0.30% points; -0.63, 0.02); p=0.06, moderate certainty). Nevertheless, following stratification by subgroups, TRE resulted in a reduction in HbA1c of 0.93 mmol/mol (-1.70, -0.17 [-0.09% points; -0.16, -0.02]; p=0.02) in individuals with prediabetes but not in individuals with type 2 diabetes (-4.68 mmol/mol; -10.08, 0.72 (-0.43% points; -0.92, 0.07); p=0.09). TRE reduced fasting blood glucose in the pooled analysis (-0.30 mmol/l; -0.53, -0.07; p<0.01, moderate certainty) as well as in the subgroup analyses in individuals with prediabetes (-0.14 mmol/l; -0.27, -0.01; p=0.03) and with type 2 diabetes (-0.48 mmol/l; -0.78, -0.17; p<0.01). Moreover, TRE lowered body weight by 1.6 kg (-2.2, -1.0; p<0.001) in the pooled analysis. The evidence was limited by imprecision arising from wide confidence intervals in some of the included studies, which may be due to small sample sizes. Lastly, the effects of TRE on markers of insulin sensitivity, beta cell function and continuous glucose monitoring measurements were inconclusive. CONCLUSIONS/INTERPRETATION: Moderate-certainty evidence indicates that TRE reduces fasting blood glucose but not HbA1c. The subgroup analyses revealed that TRE improved HbA1c and fasting glucose in individuals with prediabetes and improved fasting glucose in individuals with type 2 diabetes. Future large-scale studies should investigate long-term effects of TRE in prevention and treatment of type 2 diabetes. TRIAL REGISTRATION: PROSPERO CRD42024523591 FUNDING: This research received no specific grant from any funding agency in the public, commercial or not-for-profit sectors. Three authors (JS, A-DT, THA) are employed at Steno Diabetes Center Copenhagen, a public hospital and research institution under the Capital Region of Denmark, partly funded by a grant from the Novo Nordisk Foundation.

Humans

Critical insights on the application of the theory of planned behaviour to food handlers' food safety practices.

Foodborne diseases remain a significant public health concern, often linked to unsafe food-handling practices. The Theory of Planned Behaviour (TPB) is widely used to predict and explain food safety behaviours, yet its application in this field has not been systematically and in-depth evaluated. This review evaluated how the TPB has been applied to study food handlers' behaviour, focusing on methodological approaches, use of the TACT (Target, Action, Context, and Time) framework, validity, elicitation studies, and reliability. Seventeen studies were included following a systematic search of four databases (Scopus, Web of Science, Wiley Online Library, and Taylor & Francis Online). Data were extracted on behaviour definition, aim of study, main findings, use of indirect and direct TPB measures, use of elicitation studies, internal consistency, content validation, analytical methods used, and any extensions to the original TPB framework. Key elements related to adherence to core TPB principles and measurement practices were extracted using a Checklist. Most studies used direct measures of TPB constructs, and only a few reported procedures for content validation. Considerable variability was found in the reporting of key measurement and psychometric practices. Five studies fully applied the TACT framework, while nine incorporated additional factors such as knowledge and moral norms. Elicitation studies were conducted in five cases where indirect measures were employed. Analytical approaches were mainly based on multiple linear regression, with limited use of more advanced techniques such as structural equation modeling. Twelve studies reported internal consistency results. Overall, the review highlights opportunities to strengthen methodological practices in future TPB research on food safety. Greater attention to conducting and reporting content validation, full application of the TACT framework, reporting of internal consistency, and consistent inclusion of elicitation studies when using indirect measures may enhance transparency, reinforcing the credibility and trustworthiness of research findings. A major methodological limitation of this review was that screening and data extraction were conducted by a single reviewer and no formal quality or risk-of-bias assessment of the included studies was performed. Despite these limitations, the findings provide practical guidance for the development and validation of TPB-based questionnaires and may support more robust food safety research, interventions, and policy initiatives aimed at improving food handlers' practices.

Humans

Efficacy, acceptability, and related outcomes of pharmacological interventions for acute bipolar mania: a systematic review and dose-related network meta-analysis across different age groups.

BACKGROUND: Acute bipolar mania carries negative social and economic consequences. We investigated the comparative efficacy/response/acceptability of pharmacological interventions for acute bipolar mania, considering dose effects across different age groups. METHODS: We conducted a network meta-analysis (NMA) to search for randomized controlled trials (RCTs) comparing pharmacological interventions with one another or placebo in acute bipolar mania patients, indexed in PubMed/MEDLINE, Embase, Web of Science, and Scopus (from inception through 2025.12.24). Co-primary outcomes were change in manic symptoms/response/and acceptability. Tolerability/remission and rate of adverse events were secondary outcomes. Confidence-In-Network-Meta-Analysis was likewise appraised. RESULTS: 113 RCTs, encompassing 49 distinct treatment combinations, included 20,666 participants. Sensitivity analysis retaining only low-risk-of-bias studies and excluding outliers for possible effect modifiers indicated that risperidone 3&#xa0;mg/day(SMD&#xa0;=&#xa0;-7.57;95%C.I.&#xa0;=&#xa0;-8.25;-5.85); tamoxifen 160&#xa0;mg/day(SMD&#xa0;=&#xa0;-1.73;95%C.I.&#xa0;=&#xa0;-2.32;-1.13); rivastigmine 3&#xa0;mg/day(SMD&#xa0;=&#xa0;-1.13;95%C.I.&#xa0;=&#xa0;-1.06;-0.58); haloperidol 30&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.96;95%C.I.&#xa0;=&#xa0;-1.25;-0.75); valproate 750&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.76;95%C.I.&#xa0;=&#xa0;-1.48;-0.58); tamoxifen 40&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.75;95%C.I.&#xa0;=&#xa0;-1.41;-0.59); celecoxib 400&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.74;95%C.I.&#xa0;=&#xa0;-1.20;-0.38); paliperidone extended-release 12&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.62; 95%C.I.&#xa0;=&#xa0;-0.91;-0.32); olanzapine 15&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.59;95%C.I.&#xa0;=&#xa0;-0.60;-0.38); olanzapine 20&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.52;95%C.I.&#xa0;=&#xa0;-0.66;-0.38); risperidone 4&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.53;95%C.I.&#xa0;=&#xa0;-0.76;-0.29); allopurinol 600&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.54;95%C.I.&#xa0;=&#xa0;-0.67;-0.22); cariprazine 12&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.49;95%C.I.&#xa0;=&#xa0;-0.66;-0.33); risperidone 4.2&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.46;95%C.I.&#xa0;=&#xa0;-0.75;-0.17); lithium 1500&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.42;95%C.I.&#xa0;=&#xa0;-0.57;-0.28); ziprasidone 160&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.49;95%C.I.&#xa0;=&#xa0;-0.68;-0.31); asenapine 20&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.38;95%C.I.&#xa0;=&#xa0;-0.53;-0.22); haloperidol 8&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.34;95%C.I.&#xa0;=&#xa0;-0.63;-0.05); aripiprazole 15&#xa0;mg/day(SMD&#xa0;=&#xa0;-0.33;95%C.I.&#xa0;=&#xa0;-0.61;-0.06) outperformed placebo. Ziprasidone 160&#xa0;mg/day, celecoxib 200&#xa0;mg/day, asenapine 20&#xa0;mg/day, and asenapine 10&#xa0;mg/day proved more efficacious than placebo in children. No statistically significant differences were reported between treatments and placebo for response/remission/acceptability/tolerability, and manic/hypomanic switch. A meta-regression of efficacy effect sizes against the adapted AMSTAR-Plus content scores showed that larger SMDs were associated with lower AMSTAR scores, indicating lower study quality, warranting further caution for such large efficacy estimates. CONCLUSIONS: Our findings are consistent with previous NMAs and current guidelines, expanding the current knowledge base while concurrently appraising different drugs, doses, and age groups.

Humans

Data-centric, robust, and explainable multimodal deep learning for clinical decision support: A systematic review.

PURPOSE: Multimodal deep learning is increasingly proposed for clinical decision support (CDS) under a "data-centric" framing that prioritizes label quality, missing-modality robustness, distribution shift, calibration, and explainability. Prior reviews have examined multimodal medical AI, CDS, and data-centric methods separately, but none address their intersection. We mapped the modalities, fusion strategies, and data-centric and explainability techniques used in this recent literature, quantified how often each is implemented rather than merely mentioned, assessed deployment-relevant evidence (external validation, clinical-outcome measurement, equity), and formally appraised study-level risk of bias. METHODS: Following the PRISMA 2020 statement (PROSPERO CRD420261427815; registered retrospectively), we screened 150 records and included primary, clinical, multimodal studies that applied machine or deep learning to a decision-support task and reported at least one quantitative result. Two reviewers screened and extracted data with consensus adjudication. Each study was coded against pre-specified operational definitions, separating implemented or empirically evaluated techniques from those only mentioned. Study-level risk of bias was assessed with PROBAST + AI. Synthesis was narrative. RESULTS: Thirty-one studies met inclusion; 30 (97%) were published between 2024 and 2026, with a median of three modalities (range 2-6), most commonly structured EHR (71%) and imaging (39%). Data-centric techniques were frequently reported (74-84% across label-noise, distribution-shift, calibration, missing-modality and class-imbalance handling; equity 61%). However, external validation was reported in only 4/31 studies (13%), a clinical or provider outcome in 3/31 (10%), and no study reported routine deployment. Overall risk of bias was high in 27/31 studies (87%), driven by the analysis domain. CONCLUSION: Within this recent, self-selected slice of the field, technical robustness and explainability techniques are widely reported but rarely validated out-of-distribution or against clinical outcomes, and the underlying evidence is at high risk of bias. Progress requires external multi-site validation, clinical-outcome measurement, formal bias appraisal, and adherence to AI reporting standards (e.g., TRIPOD + AI) before deployment can be justified.

Deep Learning

Vitamin B12 Deficiency in Sickle Cell Disease: Method-Driven Estimates and Systematic Diagnostic Misclassification.

OBJECTIVES: To determine whether the reported 0%-70% prevalence of vitamin B12 deficiency in sickle cell disease (SCD) reflects true population variation or diagnostic misclassification. METHODS: We conducted a PRISMA 2020-compliant systematic review of observational studies (January 1, 2000-May 13, 2026; PROSPERO CRD420251087800) assessing B12 status in SCD. PubMed, AJOL, and Google Scholar were searched with citation tracking and dual screening. Diagnostic validity was assessed across biomarker strategy, analytical platform, thresholds, and confounder control using a proposed context-integrated framework to classify methodological robustness and discordance. RESULTS: Fourteen studies were included (57% high-income; 43% LMIC). The evidence base was dominated by limited diagnostic approaches: 71% used immunoassays, over one-third relied on circulating B12 alone, and functional biomarkers were inconsistently applied without systematic confounder adjustment. Prevalence estimates were strongly influenced by diagnostic methods rather than underlying population biology, ranging from 0% to 70% in single-marker studies (mostly 0%-7.1%, with outliers ~50%-70%) and 6.9%-53% in multi-marker studies. Discordance was substantial and greater in LMIC settings than HIC. CONCLUSION: Current diagnostic approaches in SCD appear method-dependent, generating heterogeneous prevalence estimates with uncertain clinical validity. These findings challenge existing estimates and have implications for clinical practice, research design, and diagnostic equity. TRIAL REGISTRATION: ClinicalTrials.gov identifier: CRD420251087800.

Humans

Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.

BACKGROUND: Predictive modelling for cancer risk, treatment-related complications, and survival is central to precision oncology. Conventional logistic regression (LR) and Cox proportional hazards (CoxPH) regression remain widely used but are limited when modelling nonlinear interactions, high-dimensional imaging features, and multimodal clinical-metabolic predictors. Artificial intelligence (AI) and machine learning (ML) methods offer expanded capability through automated feature extraction, ensemble learning, and flexible survival modelling, but the evidence on when AI/ML adds value over conventional models across cancer sites and predictive tasks remains fragmented. OBJECTIVE: To systematically evaluate the methodological performance, validation strategies, and translational limitations of AI/ML models compared with conventional statistical models in published predictive-modelling studies for breast, colorectal, or pancreatic cancer. METHODS: PubMed, Scopus, and Web of Science were searched for studies published between January 2019 and March 2025. Two reviewers independently conducted title-and-abstract screening, full-text eligibility assessment, and PROBAST risk-of-bias assessment. Sixty-five studies (n&#xa0;=&#xa0;907,567 participants) were narratively synthesised by cancer site, predictive task, model family, comparator, validation strategy, predictor modality, and calibration or explainability reporting. RESULTS: The 65 studies comprised breast cancer (n&#xa0;=&#xa0;35), colorectal cancer (n&#xa0;=&#xa0;21), and pancreatic cancer (n&#xa0;=&#xa0;9). AI/ML superiority over LR and CoxPH was task- and data-dependent. CNN- and U-Net-based models predominated in imaging and body-composition tasks, tree-based ensembles consistently outperformed LR for tabular perioperative complication prediction, and CoxPH remained competitive, and in the largest pancreatic risk study, superior to XGBoost (C-index 0.802 vs 0.723) in well-structured datasets. PROBAST analysis-domain risk was moderate in 54 of 65 studies (83%), driven by limited external validation, sparse calibration reporting (11/65), and few decision-curve analyses (7/65). CONCLUSION: AI/ML adds the most methodological value in imaging-derived feature extraction and nonlinear perioperative prediction, while conventional regression remains preferable in large, structured datasets with linear predictors. Clinical translation requires standardised body-composition definitions, external validation, calibration assessment, decision-curve analysis, and explainability, in line with TRIPOD+AI and CLAIM standards.

Humans

Specific Instruments for Caregiving Competence Among Family Caregivers of Cancer Patients: A COSMIN Systematic Review of Psychometric Properties.

OBJECTIVE: To evaluate and summarize the psychometric properties of specific instruments for caregiving competence among family caregivers of cancer patients. METHODS: Systematically searched eight databases for studies published up to November 2025. The methodological quality and psychometric properties of the instruments were evaluated using COSMIN 2.0. Evidence grades were rated using the modified GRADE system (four grades: "High," "Moderate," "Low," and "Very Low"), and recommendations were formulated (Category A: recommended, Category B: potential with further validation, and Category C: not recommended). RESULTS: Seven studies were included, comprising three specific instruments: the Care Competency Scale for Family Caregivers in Home Palliative Care (CCSHPC) (n = 1), the Caregiver Caregiving Self-Efficacy Scale-Oral Cancer (CSES-OC) (n = 1), and the Caring Ability of Family Caregivers of Patients with Cancer Scale (CAFCPCS) (n = 5). Both the CCSHPC and CAFCPCS received Category B recommendations, demonstrating "adequate" content validity with evidence grades rated "very low" and "low," respectively. The CAFCPCS also shows good structural validity ("moderate") and internal consistency ("low") in some cultural contexts. The CSES-OC is a Category C recommendation, with high-quality evidence indicating "inadequate" criterion validity. CONCLUSION: Few specific instruments exist, and most did not strictly follow COSMIN guidelines. The CAFCPCS is provisionally recommended based on relative evidence superiority rather than complete psychometric validation. Further cross-cultural and localized instrument development is warranted. IMPLICATIONS FOR NURSING PRACTICE: Use well-validated specific instruments to identify strengths and weaknesses in the caregiving competencies of family caregivers of cancer patients, enabling them to deliver high-quality home-based cancer care.

Female

Ibuprofen versus acetaminophen for acute mild-to-moderate pain management in pediatric populations: a systematic review and meta-analysis of their efficacy.

UNLABELLED: Ibuprofen and acetaminophen are the most widely used analgesics in pediatric practice for the management of acute mild-to-moderate pain. Despite their widespread use, the comparative analgesic efficacy of these two agents in children remains a subject of ongoing debate, with existing evidence largely derived from heterogeneous clinical settings and small individual trials. Therefore, this study aimed to systematically review and meta-analyze randomized controlled trials comparing the analgesic efficacy of ibuprofen versus acetaminophen in pediatric populations with acute mild-to-moderate pain. A systematic literature search was conducted up to May 2026 in PubMed, Scopus, and Web of Science. The review was conducted and reported in accordance with the PRISMA-Children and Adolescents (PRISMA-C) 2026 reporting guideline. Eligible studies were randomized controlled trials comparing ibuprofen with acetaminophen in children and adolescents (defined as individuals aged 0 to&#x2009;<&#x2009;18&#xa0;years) with acute pain, reporting at least one extractable efficacy outcome. Continuous outcomes were synthesized as standardized mean differences (Hedges' g) using random-effects models; dichotomous outcomes were pooled as risk ratios (RRs) with 95% confidence intervals. Risk of bias was assessed using the Cochrane RoB 2 tool and certainty of evidence was evaluated using the GRADE framework. Eight randomized controlled trials enrolling 1325 participants were included. Three pediatric trials contributed to the primary continuous pain outcome meta-analysis (n&#x2009;=&#x2009;196 analyzable participants), yielding a pooled SMD of&#x2009;-&#x2009;0.28 (95% CI&#x2009;-&#x2009;0.57 to 0.00; p&#x2009;=&#x2009;0.052; I2&#x2009;=&#x2009;0%), indicating a small effect favoring ibuprofen that did not reach conventional statistical significance. Given the small number of contributing studies (k&#x2009;=&#x2009;3), the I2 statistic should be interpreted with caution as it has limited power to detect heterogeneity in this context. For the dichotomous pain freedom outcome (2 trials, n&#x2009;=&#x2009;114), no significant difference was observed (pooled RR 1.03, 95% CI 0.53-1.99; p&#x2009;=&#x2009;0.93; I2&#x2009;=&#x2009;0%). A prespecified sensitivity analysis including an adult soft-tissue injury trial attenuated the pooled effect toward the null (SMD&#x2009;-&#x2009;0.15, 95% CI&#x2009;-&#x2009;0.38 to 0.09; p&#x2009;=&#x2009;0.23; I2&#x2009;=&#x2009;36.6%). Narrative synthesis of additional studies generally demonstrated comparable analgesic efficacy between the two agents across postoperative and outpatient pediatric settings. The overall certainty of evidence was rated as low for both primary outcomes, primarily due to imprecision and indirectness. CONCLUSION: Current evidence from randomized controlled trials does not demonstrate a superiority of ibuprofen over acetaminophen for acute mild-to-moderate pain management in children. Both agents appear to provide clinically meaningful analgesia across heterogeneous pediatric pain settings. The clinical choice between agents should be guided by individual patient factors, including contraindications to NSAIDs, the inflammatory nature of the pain etiology, and patient-specific characteristics. The low certainty of evidence underscores the need for adequately powered, methodologically rigorous trials to definitively establish the comparative efficacy of these two analgesics in the pediatric population. WHAT IS KNOWN: &#x2022; Ibuprofen and acetaminophen are the two most widely used non-opioid analgesics for acute mild-to-moderate pain in children, and both are recommended as first-line agents by major international guidelines. &#x2022; Prior meta-analyses in mixed pediatric-adult populations have suggested a modest analgesic advantage of ibuprofen over acetaminophen, but pediatric-specific evidence has remained limited and methodologically heterogeneous. WHAT IS NEW: &#x2022; This systematic review and meta-analysis, restricted to randomized controlled trials in pediatric populations, found that ibuprofen showed a small effect favoring pain reduction compared with acetaminophen (SMD&#x2009;-&#x2009;0.28, p&#x2009;=&#x2009;0.052), although this did not reach conventional statistical significance. &#x2022; The analgesic advantage of ibuprofen may be more pronounced in pain etiologies with a significant inflammatory component (e.g., fractures). At the same time, both agents appear broadly equivalent in most other acute pediatric pain settings, supporting individualized analgesic selection based on clinical context and patient-specific factors.

Humans

Diagnostic performance of machine learning models for malignant and non-malignant pleural effusion: Systematic review and meta-analysis.

BACKGROUND: Accurately distinguishing malignant pleural effusion (MPE) from non-malignant pleural effusion is clinically important, but the generalisability and methodological quality of machine-learning (ML) models remain uncertain. METHODS: We searched eight databases to 23 April 2026. Diagnostic performance was pooled using random-effects and Reitsma bivariate models, and study quality was assessed using PROBAST+AI. RESULTS: Forty-two studies were included; 17 contributed to the AUC meta-analysis and 14 to the bivariate analysis. The pooled AUC was 0.90 (95&#xa0;% CI 0.85-0.94; 95&#xa0;% prediction interval 0.62-0.98), with sensitivity of 0.80 (95&#xa0;% CI 0.77-0.83) and specificity of 0.87 (95&#xa0;% CI 0.79-0.92). Only nine studies reported external, temporal or independent validation. Externally validated studies had a lower pooled AUC than studies without external validation (0.83 vs 0.92), with lower specificity observed in the two externally validated studies contributing sensitivity and specificity data. All 42 development assessments had high overall quality concerns, and all 42 model evaluations were judged at high risk of bias. CONCLUSIONS: ML models showed good apparent accuracy for distinguishing MPE from non-MPE, but the evidence was limited by substantial heterogeneity, high risk of bias and scarce external validation. The pooled estimates reflect the average performance of different selected models rather than the expected accuracy of a single clinical test. ML models should be regarded as adjuncts to existing diagnostic pathways until they are confirmed by rigorous multicentre prospective external validation and clinical-impact studies.

Humans