Search PubMedSearch

SEARCH · Search PubMed

Results for “Environmental contamination assessment”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,481 records · Page 22Linked to original sources

Artificial Intelligence for Diagnosis, Risk Stratification, and Prognosis of Neuroblastoma - A Systematic Review and Meta-Analysis.

PURPOSE: To synthesizes evidence on artificial intelligence (AI) performance in neuroblastoma (NB) diagnosis, risk stratification, prognosis, and genomic characterization. MATERIALS AND METHODS: A systematic review and meta-analysis was conducted following PRISMA 2020 guidelines (PROSPERO: CRD42024539475) across five databases. Meta-analyses used random-effects models with logit-transformed Area Under the Curve (AUCs) and cluster-robust standard errors. AI models were classified as Machine Learning Models (MLM) or Hybrid Nomograms (HN) based on their construction methodology. RESULTS: Of 3,742 articles identified, 53 were included. MLMs demonstrated higher point estimates than radiologists in differential diagnosis (AUC: 0.87 vs. 0.83), though this difference was not statistically significant and carried substantial uncertainty. HNs achieved stronger performance in risk stratification (AUC: 0.87). AI-derived nomograms (AUC: 0.9) and gene signatures (AUC: 0.8) outperformed conventional prognostic markers descriptively. Chemotherapy response prediction remained below clinical utility thresholds across all model types. Only 33.9% of models reported calibration and 24.5% underwent external validation. CONCLUSIONS: AI demonstrates proof-of-concept across multiple NB clinical domains. However, clinical adoption remains premature given persistent gaps in external validation, calibration, dataset size, and pediatric-specific model development. Future studies should test these models prospectively in multicenter pediatric cohorts, ideally through COG or SIOPEN, using shared definitions for diagnosis, risk group, treatment response, and survival outcomes.

Humans

Machine learning vs. traditional methods for predicting postoperative cardiac complications after non-cardiac surgery: a systematic review and Bayesian network meta-analysis.

INTRODUCTION: Accurate prediction of peri-operative cardiac complications is critical to optimise pre-operative decision-making. Traditional risk prediction scores, such as the Revised Cardiac Risk Index, show only modest discrimination. Machine learning can model complex, non-linear relationships but their predictive performance compared with traditional scores remains unclear. METHODS: We performed a systematic review and Bayesian network meta-analysis. The primary outcome was postoperative adverse cardiac events following non-cardiac surgery. Prediction models were assessed relative to the Revised Cardiac Risk Index. As many studies evaluated multiple versions of each model type, the highest performing ('best version') and lowest performing ('worst version') results were analysed. Models were ranked using the surface under the cumulative ranking curve (SUCRA). RESULTS: Thirteen studies evaluating 54 models and 927,113 patients were included. Machine learning approaches generally outperformed traditional risk scores. Automated machine learning ranked highest (SUCRA 96.6) showed the greatest improvement in the best version analysis (mean difference (MD) 0.28 (95%CrI 0.16-0.40)) and remained superior in the sensitivity analysis (MD 0.30 (95%CrI 0.14-0.45)). Gradient boosting models showed superior performance over the Revised Cardiac Risk Index across analysis (best version: MD 0.20 (95%CrI 0.14-0.26), worst version: MD 0.18 (95%CrI 0.12-0.25), SUCRA 82.4). The Gupta Perioperative Risk for Myocardial Infarction or Cardiac Arrest score outperformed the Revised Cardiac Risk Index in the best version analysis (MD 0.16 (95%CrI 0.01-0.32)). Between-study heterogeneity was low. None of the included studies externally validated their machine learning models and only six were judged to be at low risk of bias. DISCUSSION: Most machine learning models showed better discrimination than traditional risk scores, with automated machine learning and gradient boosting models ranking highest. However, study quality, calibration reporting and absence of external validation limit immediate clinical adoption. Prospective, multicentre evaluation is required before integration of these models into peri-operative practice.

Humans

Predicting ACL injury risk in athletes: A systematic review of machine learning-based models.

BACKGROUND: Early ACL injury risk identification in athletes is essential. This systematic review examines machine learning (ML) models for predicting ACL injuries, evaluating their methodological quality, performance, and reliability. METHOD: A comprehensive electronic search was conducted across PubMed, Scopus, Web of Science, and IEEE Xplore databases, supplemented by Google Scholar for grey literature, covering articles published between January 1, 2015, and August 30, 2025. Eligible studies were appraised using the Prediction Model Study Risk of Bias Assessment Tool (PROBAST) for methodological quality and risk of bias, and the Transparent Reporting of a Multivariable Prediction Model for Individual Prognosis or Diagnosis (TRIPOD) guidelines for quality of evidence. RESULTS: Ten studies were included. PROBAST showed eight studies had moderate risk of bias and two low risk. TRIPOD found only two studies met quality criteria. ML models included logistic regression (n = 5), support vector machines (n = 4), k-nearest neighbor (n = 3), decision trees (n = 3), random forests (n = 5), neural networks (n = 2), linear discriminant analysis (n = 1), and pre-trained CNNs (n = 1). AUC ranged from 0.63 to 0.98. Accuracy (reported in six studies) ranged from 26% to 95%; however, these values should be interpreted with caution due to the absence of confidence intervals, lack of class imbalance handling, and limited external validation across studies. Tree-based ensemble methods such as random forest achieved competitive accuracy (74-86%), while SVM, a non-ensemble classifier, reported accuracy ranging from 71% to 95%; however, the highest values were obtained in studies with notably small sample sizes (n = 12 to n = 39), raising concerns about overfitting and generalizability. CONCLUSION: Current ML algorithms show promise for identifying athletes at high ACL injury risk and detecting relevant risk factors. Although study quality was generally satisfactory, future research should prioritize external validation and model interpretability to support clinical translation.

Humans

Applications and outcomes of virtual reality in inpatient psychiatry: A systematic review.

BACKGROUND: Virtual reality (VR) has been widely used in outpatient psychiatric services and has demonstrated benefits across several clinical diagnoses, but its use and effects in inpatient settings remain to be explored. This systematic review aimed to examine the use of VR during psychiatric hospitalization, including types of VR applications, barriers and facilitators of implementation, and effects on various outcomes. METHODS: The review was registered in PROSPERO (#CRD42023446524). Following PRISMA guidelines, databases (Ovid, SciVerse, Web of Science, Cochrane Library, ProQuest, and WorldCat) were searched from 1983 to 2025 using keywords related to VR and psychiatric disorders. Studies involving the use of VR with psychiatric inpatients (≥85%) were included. Descriptive statistics and narrative syntheses were used to summarize findings. Study quality was assessed with the Mixed Methods Appraisal Tool. RESULTS: After full-text screening, 37 studies (N = 1,004) met inclusion criteria. VR was used for both assessment and intervention, with cognitive-behavioral therapy/exposure (35%) and assessment (24%) being the most frequently used. VR use in inpatient units appeared feasible, acceptable, and safe for inpatients and clinicians, though findings remain preliminary. Several facilitators (e.g. adequate staff training and supervision) and common barriers (e.g. technical difficulties and limited resources) were identified. The most consistent improvements were observed in clinical symptoms (e.g. anxiety) compared with psychosocial, cognitive, and physiological outcomes. CONCLUSIONS: These findings suggest that inpatient settings represent a promising, yet understudied context for VR-based assessments and interventions. High-quality trials and systematic reporting of implementation are needed in future studies to inform research and clinical practice.

Humans

Efficacy and Safety of Autologous Versus Prosthetic Grafts in the Repair of Popliteal Artery Aneurysms: A Systematic Review and Meta-Analysis.

BACKGROUND: Popliteal artery aneurysms (PAAs) present a severe risk of progression to acute limb ischemia. Open surgery (OS) is the gold standard treatment; however, prosthetic grafts are acceptable in highly selected cases, especially when the great saphenous vein is not available. METHODS: We performed a systematic review and meta-analysis of studies comparing autologous versus prosthetic grafts for patency and limb preservation outcomes in patients with PAAs. MEDLINE, Embase, and Cochrane Central were systematically searched from inception through October 2024. Outcomes were pooled using a frequentist random-effects model as odds ratios, mean differences, and hazard ratios (HRs) with 95% confidence intervals (CIs) on RStudio (Version 4.5.0). Risk-of-bias assessments were performed using ROBINS-I and MINORS. RESULTS: Twenty-two observational studies were pooled comprising 9,145 PAAs in 8,370 patients, of whom 6,434 (74.51%) were treated with autologous grafts and 2,200 (25.49%) with prosthetic grafts. Follow-up ranged from 12 to 86 months. Repair with autologous conduits significantly improved long-term primary patency (HR 3.93; P < 0.001), secondary patency (HR 6.02; P < 0.001), and long-term limb salvage (HR 2.69; P = 0.044) compared with prosthetic conduits. There were no significant differences in in-hospital amputation (P = 0.36), myocardial infarction (P = 0.61), mortality (P = 0.50), 2-year primary patency (P = 0.25), 5-year secondary patency (P = 0.06), or length of hospital stay (P = 0.95). Risk of bias was classified as moderate-to-high, reflecting confounding factors inherent to observational studies and moderate methodological quality by MINORS. Despite these limitations, treatment effects consistently favored autologous grafts in both short- and long-term analyses; however, caution is warranted given the limited number of available studies. CONCLUSION: The use of autologous conduits significantly favors both short-term and long-term efficacy and safety in the OS repair of PAAs. Given the limitations of the existing evidence, further comparative studies are needed.

Humans

Diagnostic performance of panfungal PCR on tissue specimens for the diagnosis of invasive fungal diseases: a systematic review and meta-analysis of the Fungal PCR Initiative (FPCRI).

UNLABELLED: Invasive fungal diseases are difficult to diagnose because of the limited sensitivity of culture. Panfungal PCR amplicon sequencing assays (targeting ribosomal RNA, such as 18S, 28S, ITS) are recommended for fungal identification in histopathology samples showing fungal elements. However, data describing its overall performance and consistency are lacking. This systematic literature review and meta-analysis assessed the performance of panfungal PCR on formalin-fixed paraffin-embedded (FFPE) and non-fixed (fresh or frozen) tissue samples. A systematic literature search was performed to include studies reporting the use of panfungal PCR for fungal identification in FFPE or non-fixed tissue samples. PCR sensitivity and specificity were assessed using the reference standard of histopathology showing fungal elements. Quality assessment was performed using the Quality Assessment of Diagnostic Accuracy Studies (QUADAS-2) tool. Pooled estimates were obtained using random-effects meta-analysis. Twenty-eight studies were included. In FFPE samples (18 studies, 852 samples), sensitivity and specificity were 75.4% (95% confidence interval [CI], 59.2-86.6) and 93.5% (70.2-98.9), respectively. Sensitivity in non-fixed samples (13 studies, 207 samples) was 86.5% (74.7-93.3), while specificity could not be assessed (insufficient data). Comparative analyses showed a significantly higher sensitivity of panfungal PCR over culture (88.2%; 76-94.7 vs 52.2%; 39-65, P = 0.001). Sub-analyses could not demonstrate the superiority of one PCR target over another due to limited data. Panfungal PCR exhibited adequate sensitivity and good specificity in FFPE samples. Sensitivity was even higher in non-fixed samples and largely superior to culture. Nevertheless, large interstudy variability was observed, warranting interlaboratory studies to define the optimal PCR target and standardized protocols. IMPORTANCE: Invasive fungal diseases are difficult to diagnose because of the low sensitivity of culture. Panfungal PCRs are widely used for fungal identification in tissue specimens but suffer from heterogeneous procedures and performance. This meta-analysis shows an acceptable sensitivity (75.4% and 86.5% in fixed and non-fixed samples, respectively) and good specificity (93.5%) of panfungal PCR, supporting its use, not only on histopathology-positive fixed samples but also in non-fixed samples concomitantly with other diagnostic tools (cultures and fungal-specific PCRs if available). These results provide a strong basis for further standardization of panfungal PCR techniques via interlaboratory assays to assess reproducibility and optimize analytical protocols. CLINICAL TRIALS: This study is registered with PROSPERO as CRD42023461148.

Humans

Timing of OMERACT core domain measurement in gout clinical trials: a systematic review of randomised trials.

AIMS: The Outcome Measures in Rheumatology (OMERACT) initiative has endorsed core domain sets for gout trials. The aims of this study were to evaluate the time points and frequencies at which the gout core domains are measured in existing gout urate-lowering therapy and gout flare trials, and whether all collected measurements were reported. METHODS: Urate-lowering therapy (n = 29) and gout flare randomised clinical trials (n = 14) from 2005 were identified from a prior systematic review of core domain reporting. Data were extracted for the time points and frequencies at which each core domain was measured, as well as whether all collected measurements were reported. RESULTS: In urate-lowering therapy trials, the core domains were measured at seven different frequencies. Serum urate and gout flares were most commonly measured monthly, and tophus burden was most commonly measured three monthly. Reporting of all collected measurements varied, from 24/29 (83%) trials for serum urate to 0/2 (0%) trials for activity limitation. In gout flare trials, core domains were measured at nine different frequencies. Pain, joint tenderness and joint swelling were most commonly measured monthly. Reporting of all collected measurements varied, from 13/14 (93%) trials for pain to 3/8 (37.5%) trials for joint tenderness. CONCLUSION: In both urate-lowering therapy and gout flare trials, there is substantial variability in when the core domains are measured, and reporting of collected measurements is inconsistent. This work provides the foundation for a consensus process to establish standardised time points and frequencies for measuring the OMERACT-endorsed gout core domains.

Gout

From premature adrenarche to adult metabolic risk and hyperandrogenism: a systematic review and meta-analysis.

CONTEXT: Idiopathic premature adrenarche (IPA) has been associated with a higher risk of metabolic and reproductive dysfunction, but long-term/adult outcomes remain incompletely known. OBJECTIVE: To assess the relationship between IPA and metabolic syndrome, as well as polycystic ovarian syndrome, in premenarcheal adolescent and adult women. METHODS: We conducted a systematic review and meta-analysis of observational studies reporting outcomes in females with IPA after menarche. Databases were searched through February 2025. Primary outcomes included body mass index (BMI), insulin resistance markers, and clinical and biochemical markers of hyperandrogenism. Data were pooled using random-effects models. The GRADE approach was applied to assess the certainty of evidence. RESULTS: A total of 21 studies comprising 635 females with IPA and 307 age-matched controls were included. Compared to controls, IPA individuals showed significantly higher BMI (mean difference: 1.4; CI: 1.0-1.9), fasting insulin, and homeostasis model assessment of insulin resistance, indicating persistent insulin resistance. Markers of hyperandrogenism, including Ferriman-Gallwey score, dehydroepiandrosterone sulfate, and Free androgenic index, were also elevated. Secondary analyses revealed higher triglycerides, lower high-density lipoprotein, increased leptin, and greater carotid intima-media thickness, supporting an early pattern of cardiometabolic risk. GRADE assessment rated most outcomes as low certainty. CONCLUSION: Women with a history of IPA are at increased risk of long-term insulin resistance and hyperandrogenism, with early signs of adverse cardiometabolic profiles. These findings support the need for long-term monitoring in this population.

Humans

Joint association of sedentary behaviour and physical activity with cardiovascular disease: a systematic review and meta-analysis.

This systematic review and meta-analysis of cohort studies aimed to synthesize existing evidence on the joint association of physical activity (PA) and sedentary behaviour (SB) with cardiovascular disease (CVD) risk among adults. We searched PubMed, EMBASE, and Cochrane for English studies published between January 2010 and February 2025 that examined the joint association of PA and SB (fatal and non-fatal) CVD among adults and pooled their results through meta-analyses using study-level data. Using findings from 17 studies, the pooled effect size for the lowest PA + highest SB group was 1.78 (95% CI: 1.59-2.00), suggesting increased risk for CVD compared with the reference group (i.e. highest PA + lowest SB). Compared with the same reference group, we found an increased risk for CVD in the lowest PA + lowest SB (HR = 1.25, 95% CI: 1.12-1.40) and the highest PA + highest SB (HR = 1.16, 95% CI: 1.06-1.27) groups. Subgroup analyses according to domains or types of PA and SB exposure, outcome measures, and exposure measurement method revealed a similar pattern. In conclusion, individuals with the lowest PA combined with the highest SB may experience an increased risk for CVD events compared with those with the highest PA and lowest SB. There was also an increased risk for individuals with a combination of either low PA + low SB and high PA + high SB, albeit to a lesser extent. Although substantial study-level heterogeneity exists, the results highlight the potential value of considering both behaviours jointly in relation to CVD risk.

Humans

Critical insights on the application of the theory of planned behaviour to food handlers' food safety practices.

Foodborne diseases remain a significant public health concern, often linked to unsafe food-handling practices. The Theory of Planned Behaviour (TPB) is widely used to predict and explain food safety behaviours, yet its application in this field has not been systematically and in-depth evaluated. This review evaluated how the TPB has been applied to study food handlers' behaviour, focusing on methodological approaches, use of the TACT (Target, Action, Context, and Time) framework, validity, elicitation studies, and reliability. Seventeen studies were included following a systematic search of four databases (Scopus, Web of Science, Wiley Online Library, and Taylor & Francis Online). Data were extracted on behaviour definition, aim of study, main findings, use of indirect and direct TPB measures, use of elicitation studies, internal consistency, content validation, analytical methods used, and any extensions to the original TPB framework. Key elements related to adherence to core TPB principles and measurement practices were extracted using a Checklist. Most studies used direct measures of TPB constructs, and only a few reported procedures for content validation. Considerable variability was found in the reporting of key measurement and psychometric practices. Five studies fully applied the TACT framework, while nine incorporated additional factors such as knowledge and moral norms. Elicitation studies were conducted in five cases where indirect measures were employed. Analytical approaches were mainly based on multiple linear regression, with limited use of more advanced techniques such as structural equation modeling. Twelve studies reported internal consistency results. Overall, the review highlights opportunities to strengthen methodological practices in future TPB research on food safety. Greater attention to conducting and reporting content validation, full application of the TACT framework, reporting of internal consistency, and consistent inclusion of elicitation studies when using indirect measures may enhance transparency, reinforcing the credibility and trustworthiness of research findings. A major methodological limitation of this review was that screening and data extraction were conducted by a single reviewer and no formal quality or risk-of-bias assessment of the included studies was performed. Despite these limitations, the findings provide practical guidance for the development and validation of TPB-based questionnaires and may support more robust food safety research, interventions, and policy initiatives aimed at improving food handlers' practices.

Humans

Risk of stroke in SLE: a systematic review and meta-analysis.

UNLABELLED: The association between SLE and composite stroke, ischaemic stroke and haemorrhagic stroke remains incompletely understood. This meta-analysis aims to assess the risk of stroke in patients with SLE. METHODS: Data sources included PubMed, Embase, the Cochrane Library and reference lists of included studies. This meta-analysis included cohort studies evaluating whether stroke risk is associated with SLE. The risk of bias was assessed using the Newcastle-Ottawa Quality Assessment Scale (NOS). Risk ratios (RRs) with 95% CIs were pooled using a random-effects model, and publication bias was assessed with funnel plots and Egger's test. RESULTS: A total of 25 cohort studies involving 5&#x2009;220&#x2009;837 individuals were included in this meta-analysis, which were published between 2001 and 2026. The pooled analysis demonstrated a significantly increased risk of stroke in patients with SLE (RR of 2.60, 95%&#x2009;CI 2.21 to 3.05, I&#xb2;=97.9%, p<0.001). The risk of composite stroke (RR of 2.83, 95%&#x2009;CI 2.25 to 3.57, I&#xb2;=98.0%, p<0.001), ischaemic stroke (RR of 2.34, 95%&#x2009;CI 1.75 to 3.12, I&#xb2;=97.6%, p<0.001) and haemorrhagic stroke (RR of 2.66, 95%&#x2009;CI 1.57 to 4.49, I&#xb2;=96.2%, p<0.001) was also increased in SLE. Despite the large heterogeneity, the sensitivity analysis indicated that the results were robust, and there was little evidence of publication bias. CONCLUSION: The risk of composite stroke, ischaemic stroke and haemorrhagic stroke is increased in SLE. PROSPERO REGISTRATION NUMBER: CRD420261294082.

Humans

Ultrasound protocols used to detect vascular gas emboli in divers: a systematic review.

INTRODUCTION: Venous gas emboli (VGE) detected via ultrasound can be used as a surrogate marker for decompression stress. While Doppler ultrasound is the historical gold standard, two-dimensional (2D) ultrasonography offers advantages for on-site monitoring, including a wider field of view and reduced dependence on noise-free environments. This systematic review evaluates 2D ultrasonography protocols used in decompression research since the 2015 International Meeting on Ultrasound for Diving Research, identifying methodological similarities, differences, and adherence to consensus recommendations. METHODS: A search of PubMed and Scopus identified studies using 2D ultrasound to detect VGE in divers. Inclusion criteria were: (1) use of 2D ultrasound, (2) detection of VGE or monitoring of decompression stress, (3) inclusion of a diver cohort, and (4) publication after 2015. Data extraction focused on VGE scoring systems, ultrasound hardware, measurement protocols, and operator experience. Risk of bias was assessed using ROBINS-I-V2, and compliance with the 2015 consensus recommendations was evaluated. RESULTS: Twenty studies were included. The Eftedal-Brubakk scale was most commonly used (n = 15), with cardiac ultrasound as the primary imaging modality; one study assessed a peripheral vessel. Common shortcomings included post-dive measurements lasting less than two hours, underreporting of operator experience and hardware specifications, limited individual-level data, and inappropriate use of parametric statistics for ordinal bubble grade data. No study fully complied with all consensus recommendations. CONCLUSIONS: This review demonstrates that, although two-dimensional ultrasound is widely used for post-dive VGE assessment, methodological heterogeneity with multiple shortcomings remain. Furthermore, nearly all studies restricted imaging to the heart, thus leaving peripheral vessel assessment largely unexplored.

Embolism, Air

Adjuvant CDK4/6 inhibitors in early-stage breast cancer: Clinical evidence and considerations for risk stratification and treatment selection.

Hormone receptor-positive, human epidermal growth factor receptor 2-negative breast cancer is the most common biologic subtype and carries a persistent risk of recurrence, particularly in patients with high-risk, early-stage disease. Cyclin-dependent kinase 4 and 6 inhibitors, initially established as a standard component of first-line therapy in the metastatic setting based on improvements in progression-free and overall survival, have since been evaluated in the adjuvant setting. While adjuvant palbociclib did not improve invasive disease-free survival, the monarchE and NATALEE trials demonstrated that abemaciclib and ribociclib, respectively, reduce recurrence risk in patients with high-risk, early-stage disease, with emerging overall survival data further supporting their use. However, the absolute magnitude of benefit varies substantially with baseline risk, and treatment-related toxicity and adherence challenges must be considered, as approximately 20% to 25% of patients discontinue therapy before completion. The integration of these agents into clinical practice also intersects with ongoing efforts to deescalate axillary surgery, as treatment eligibility has been largely defined by anatomic staging, particularly nodal status. Available data suggest that the incremental impact of axillary surgery on identifying candidates for cyclin-dependent kinase 4 and 6 inhibition is modest, especially among the favorable-risk populations now eligible for surgical deescalation. As the field evolves, advances in molecular risk stratification, genomic profiling, and dynamic biomarkers are poised to shift treatment selection from anatomic staging toward biologically driven approaches. Multidisciplinary decision-making that integrates tumor biology, anticipated absolute benefit, toxicity, patient preferences, and surgical considerations will be essential to ensure individualized care.

Humans

Updated adjunctive minocycline for schizophrenia: A systematic review and meta-analysis of clinical and cognitive outcomes.

BACKGROUND: Minocycline has been proposed as an adjunctive treatment for schizophrenia due to its anti-inflammatory and neuroprotective properties. However, evidence regarding its efficacy across clinical and cognitive outcomes remains inconsistent. METHODS: A systematic review and meta-analysis of double-blind RCTs was conducted following PRISMA guidelines. PubMed, Web of Science, Embase, Ovid MEDLINE, and the Cochrane Library were searched from January 2000 to August 2025. Eligible studies included patients with schizophrenia receiving adjunctive minocycline plus stable antipsychotics. Primary outcomes were PANSS total and subscale scores and overall cognitive performance. Secondary outcomes included SANS, CDS, CGI, GAF, and seven cognitive domains. Standardized mean differences (SMDs) with 95% CIs were calculated. RESULTS: Ten RCTs involving 895 participants were included. Adjunctive minocycline was associated with improvements in negative symptoms (PANSS negative: SMD = -0.55, 95% CI: -0.96 to -0.13; SANS: SMD = -0.75, 95% CI: -1.00 to -0.49) and overall psychopathology (PANSS total: SMD = -0.49, 95% CI: -0.80 to -0.18). Cognitive benefits were limited to a modest improvement in working memory (SMD = 0.24, 95% CI: 0.08 to 0.39), with no significant effects in other cognitive domains. Subgroup analyses suggested that illness stage, antipsychotic regimen, treatment duration, sample size, and geographic region may contribute to variability in treatment effects. Adverse event rates were comparable between groups. CONCLUSIONS: Adjunctive minocycline may improve negative symptoms and provide modest working memory benefits in schizophrenia. However, the evidence is limited by substantial heterogeneity, potential small-study effects, and inconsistent findings. Although short- to medium-term tolerability appeared comparable to placebo, larger, longer-term RCTs are needed to confirm its efficacy and safety.

Humans

Near and distance vergence facility provides complementary clinical information in concussion-related convergence insufficiency.

PURPOSE: To compare near and distance vergence facility testing in adolescents and young adults with concussion-related convergence insufficiency and evaluate changes following office-based vergence/accommodative therapy (OBVAM). METHODS: This secondary analysis of the CONCUSS randomized clinical trial evaluated vergence facility at near (40&#xa0;cm) and distance (4&#xa0;m) using a 12&#x394; base out/3&#x394; base in prism flipper. Participants aged 11-25&#xa0;years with concussion-related convergence insufficiency were randomized to immediate or 6&#xa0;weeks delayed OBVAM. Vergence facility was assessed at baseline, outcome time 1 assessment (after 12 therapy sessions for the immediate group and 6&#xa0;weeks of watchful waiting for the delayed group), and outcome time 2 assessment (after both groups completed 16 therapy sessions). Agreement between near and distance vergence facility classifications was evaluated, and treatment-related changes were compared between groups. RESULTS: Of the 106 enrolled participants, 102 completed all study visits. At baseline, the near and distance vergence facility classifications demonstrated substantial discordance. Among 101 participants with both measures available, 49 demonstrated reduced distance vergence facility despite normal near vergence facility, whereas only two showed the opposite pattern (Cohen's &#x3ba;&#xa0;=&#xa0;0.12; p&#xa0;<&#xa0;0.0001). Vergence facility improved following therapy in both treatment groups, with larger early improvements in the immediate-treatment group. CONCLUSIONS: Near and distance vergence facility testing provided complementary rather than interchangeable clinical information in adolescents and young adults with concussion-related convergence insufficiency. Both measures improved following vergence/accommodative therapy, supporting consideration of both testing distances in clinical assessment.

Humans

Risk-Benefit of Phase 2 Monotherapy Trials in Adult Solid Cancers: A Systematic Review and Meta-Analysis.

Phase 2 (and phase 1 dose expansion, which we label phase 2 for the purposes of this study) cancer trials are the first direct tests of a new drug's efficacy. Because efficacy evidence is lacking, the therapeutic status of drug administration during ethical review is uncertain. We compared the efficacy and safety of cancer monotherapies in phase 2&#xa0;with phase 3, where clinical equipoise underwrites a therapeutic status for drug administration. In this systematic review and meta-analysis, we searched Clinicaltrials.gov for phase 2 and 3 investigational monotherapy drug trials in six solid malignancies, with primary completion dates 2015-2020, inclusive. Two independent reviewers completed data extraction. Effects were estimated using an inverse-variance weighted random-effects model meta-analysis of proportions using the R package meta. We analyzed 130 phase 2 and 52 phase 3 trial arms, enrolling 6665 and 18,694 patients. The pooled objective response rate was 7% (95% CI 5%-11%) in phase 2 versus 24% in phase 3 (95% CI 17%-31%; p&#x2009;<&#x2009;0.0001). The median PFS and OS were shorter in phase 2 compared to phase 3 (3.23 vs. 5.43&#x2009;months, p&#x2009;<&#x2009;0.0001; 9.46 vs. 14.44&#x2009;months, p&#x2009;=&#x2009;0.0001). The pooled rate of drug-related grade 3-4 adverse events was 30% (95% CI 23%-37%) in phase 2 and 25% (95% CI 19%-32%) in phase 3. Monotherapies delivered in phase 2 cancer trials present diminished risk-benefit compared with phase 3 and align with historic estimates for phase 1. Though there may be exceptions, risks for drug administration in phase 2 should generally be justified by appeals to research rather than therapeutic value.

Humans

Diagnostic Performance of Machine Learning for Systemic Lupus Erythematosus: Systematic Review and Meta-Analysis.

BACKGROUND: Early and accurate diagnosis of systemic lupus erythematosus (SLE) and its organ involvement is essential. Previous reviews of machine learning (ML) in SLE combined heterogeneous tasks and validation strategies and may have overinterpreted model performance. OBJECTIVE: This study evaluated the diagnostic performance of ML and deep learning (DL) models for 3 clinically distinct SLE-related tasks: SLE classification or diagnosis, lupus nephritis (LN) diagnosis, and neuropsychiatric systemic lupus erythematosus (NPSLE) discrimination. We also assessed methodological quality and certainty of evidence. METHODS: PubMed, Embase, Cochrane Library, Web of Science, and IEEE Xplore were searched from January 2014 to April 2026. Eligible peer-reviewed diagnostic accuracy studies developed or validated ML or DL models for 1 of the 3 prespecified tasks, used an accepted reference standard, and provided data for a 2&#xd7;2 contingency table. Bivariate random-effects meta-analyses with the Hartung-Knapp-Sidik-Jonkman adjustment were used to pool sensitivity and specificity. We reported 95% prediction intervals (PIs), assessed risk of bias using the Quality Assessment of Diagnostic Accuracy Studies for Artificial Intelligence tool (QUADAS-AI; Viknesh Sounderajah [Imperial College London]), and evaluated certainty of evidence using the Grading of Recommendations Assessment, Development, and Evaluation framework for diagnostic test accuracy. RESULTS: Twenty-nine studies were included: 17 for SLE classification, 5 for LN diagnosis, and 7 for NPSLE discrimination. In the primary task-stratified analysis, pooled sensitivity was 0.91 (95% CI 0.86-0.94; 95% PI 0.56-0.99), and pooled specificity was 0.94 (95% CI 0.91-0.96; 95% PI 0.69-0.99), with low heterogeneity (I&#xb2;=23.9% and 22.9%, respectively). DL models showed a sensitivity of 0.93 and specificity of 0.95, compared with 0.88 and 0.94 for traditional ML models. Certainty of evidence was high for most analyses but low for LN diagnosis because of inconsistency and imprecision. All studies were retrospective, and only 9 of 29 (31%) performed independent external validation. Overall risk of bias was high or unclear in 22 of 29 (75.9%) studies. No study reported model calibration, decision-curve analysis, or net clinical benefit. CONCLUSIONS: ML models showed promising diagnostic accuracy across 3 distinct SLE-related tasks, but wide PIs, limited external validation, and pervasive risk of bias restrict conclusions about real-world generalizability. Prospective multicenter studies with standardized tasks and reference standards, independent external validation, and formal assessment of calibration and clinical utility are required before clinical implementation.

Humans

Aortoesophageal Fistula: Mending the Lethal Connection-A Systematic Review and Meta-Analysis.

BACKGROUND: Aortoesophageal fistula (AEF) is a rare, life-threatening condition with limited high-quality evidence to guide management. METHODS: We conducted a systematic review and meta-analysis of PubMed and Scopus through December 2025, evaluating survival and complication outcomes according to treatment strategy and morphological severity. Patients were categorized into 5 groups: TEVAR alone, staged TEVAR followed by surgery, primary open or hybrid surgery, TEVAR with esophageal stenting, and supportive/palliative care. Outcomes were pooled as proportions with 95% confidence intervals (CIs). RESULTS: A total of 167 reports representing 528 patients were included. Overall 30-day mortality was 31.2% (95% CI 27.0% to 35.7%) and 1-year mortality 41.2% (95% CI 36.2% to 46.4%). Infection occurred in 32.6%, reintervention in 20.2%, and recurrence in 13.2%. Staged TEVAR followed by surgery showed favorable survival (30-day 12.7%, 1-year 26.1%) but high infection (66.7%) and reintervention (43.2%) rates. TEVAR with esophageal stenting had the lowest early mortality (5.0%) but frequent reintervention (54.5%) and 1-year mortality of 33.3%. TEVAR alone demonstrated the lowest 1-year mortality among definitive strategies (25.5%) but notable recurrence (26.9%). Primary open or hybrid surgery carried higher early and late mortality, while supportive/palliative care had the worst outcomes. Morphological severity correlated strongly with mortality, infection, reintervention, and recurrence, with type IV lesions showing particularly poor prognosis. CONCLUSION: Despite contemporary management, AEF carries high early and late mortality. Definitive strategies, particularly staged TEVAR followed by surgery, offer improved survival but increased complications. Morphology-guided, individualized management is recommended.

Adult