Search PubMedSearch

SEARCH · Search PubMed

Results for “Clinical performance”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,744 records · Page 7Linked to original sources

Can ChatGPT Replace Human Clinical Coders? A Comparative Study in Otology Billing.

OBJECTIVE: Evaluate the utility of the large language model (LLM), ChatGPT, for the analysis of operative notes and the generation of Current Procedural Terminology (CPT) codes in comparison to human clinical coders. STUDY DESIGN: CPT billing codes assigned by ChatGPT were compared to existing billing data. Otology practice within a tertiary academic center. METHODS: About 191 operative notes from a single surgeon (9/2022-10/2023) were analyzed. ChatGPT-3.5 and 4 models were prompted for CPT codes based on operative notes. Assessment included determining exact and partial match rates, sensitivity and specificity for targeted procedures, and work Relative Value Units (wRVU) differences between ChatGPT-generated and human-assigned codes. RESULTS: ChatGPT-3.5 achieved exact matches in 22% of cases and partial matches in 32%, while ChatGPT-4 achieved 14% exact and 33% partial matches. When cochlear implantation (CI) was excluded, performance dropped significantly. For CI, ChatGPT-3.5 demonstrated a sensitivity of 94% and specificity of 90%, while ChatGPT-4 showed a sensitivity of 96% and specificity of 92%. In contrast, performance on cartilage grafting was poor, with sensitivities of 4.2% for ChatGPT-3.5 and 0% for ChatGPT-4. ChatGPT-3.5 and 4 showed moderate CPT code matching accuracy among themselves, with slight agreement to human coders. Both models tended to underbill for wRVUs compared to human coders, with significant differences in the values generated. CONCLUSION: This study assessed ChatGPT's effectiveness in automating CPT code assignment for otologic surgeries. While the models achieved high sensitivity values for assigning codes related to cochlear implantation, both models struggled with complex cases, failed to apply modifiers, and often assigned fewer wRVUs. The findings highlight ChatGPT's potential in medical billing but indicate a need for further refinement.

Humans

Methods for defining equity-stratifying variables: a systematic review of validation studies.

BACKGROUND AND OBJECTIVE: Disease burden is often disproportionally higher among those who are socially disadvantaged by factors defined in the PROGRESS-Plus framework (ie, Place of residence, Race/ethnicity/culture/language, Occupation, Gender/sex, Religion, Education, Socioeconomic status, and Social capital, with "Plus" covering features like age and disability). The accuracy and applicability of case definitions to identify these variables from administrative and clinical health data are unknown. We conducted a systematic review to explore how equity-stratifying variables, as categorized by the PROGRESS-Plus framework, have been defined and validated in epidemiologic studies using administrative health, population-level, or electronic health record (EHR) data. METHODS: Medline, EMBASE, CINAHL, Web of Science, and Google Scholar were searched from the inception of the databases to 2024 for validation studies of equity-stratifying variables in adults using administrative health datasets, health registries, or EHR data. Titles and abstracts, followed by relevant full-text articles, were screened in duplicate by two reviewers for eligibility. The data sources utilized, algorithms employed, and their associated performance measures were extracted and synthesized from included studies. Given substantial heterogeneity in study design, equity-stratifying variable definition, and performance metrics, meta-analysis was not possible. RESULTS: Of the 9099 unique citations screened, 188 full texts were reviewed and 116 were included in this review. Most studies were published between 2019 and 2024 (n = 64, 55%) and were validation studies of race/ethnicity definitions that used race/ethnicity codes or surname list algorithms (n = 66, 57%). No studies examined religion. Regarding the reported performance measure estimates, the race/ethnicity/culture/language equity-stratifying variables category had the largest variability across sensitivity, positive predictive value (PPV), and Cohen's Kappa. Occupation validation studies had the lowest variation in sensitivity and PPV. CONCLUSION: Despite an increasing number of publications reporting on the validation of equity-stratifying variables relevant to the PROGRESS-Plus framework, performance measures varied widely across studies. The significant heterogeneity in equity-stratifying variable definitions and methods used to validate them support the need for further rigorous validation of equity-stratifying variables in administrative and clinical health data. PLAIN LANGUAGE SUMMARY: Disease burden is often higher in people who experience financial hardships, lower level of education, discrimination due to race/ethnicity, and unstable housing. These social factors can be considered health equity factors and are important for understanding health inequalities. Health researchers often use large datasets, such as hospital or electronic health records (EHRs), to study these health equity factors. However, it is not clear how accurately these data sources capture information about people's social circumstances and how these factors are defined. In this study, we reviewed existing research to understand how health equity factors have been defined across health data sources and how accurate they are at measuring aspects of health equity and social disadvantage. Of the more than 9000 studies we identified, we included 116 that met our criteria for this systematic review. Most included studies focused on identifying race and ethnicity, often using codes or surname-based methods. We found that the accuracy of these methods varied widely across studies, meaning results may not always be reliable or comparable. Overall, our findings show that there are inconsistencies in how social factors are defined and measured in health data. This makes it difficult to fully understand and address health inequalities using routinely collected health data. More work is needed to develop and validate better quality and more consistent methods for capturing these important social factors.

Humans

Nimodipine in animal models of demyelination relevant to multiple sclerosis: a systematic review.

BACKGROUND: Multiple sclerosis (MS) is the most common inflammatory neurodegenerative disease in which axonal injury, neuronal death, and demyelination occur. Treatment for MS relapses remains limited, which alleviates acute loss of function but has no impact on long-term disability. This study aimed to perform a systematic review of the effects of nimodipine on experimental demyelination models, including experimental autoimmune encephalomyelitis (EAE) and Cuprizone models in rodents. METHODS: This study was conducted following the PRISMA statement. A systematic search was performed in PubMed, Scopus, the Cochrane Library, and Google Scholar. The primary outcome was EAE clinical disease severity (peak clinical score and/or cumulative disease burden). Secondary outcomes included relapse activity (when reported), histological myelin outcomes, oligodendrocyte lineage markers, neuroaxonal injury markers, and inflammatory readouts. Risk of bias was assessed using the SYRCLE tool. RESULTS: Out of 4660 results, 5 studies were included in the systematic review (four EAE studies and one cuprizone model). Nimodipine was administered using heterogeneous regimens (oral, intravenous, intraperitoneal, subcutaneous, or osmotic pump delivery; 1-30 mg/kg/day). The included studies reported the variable effects of nimodipine on relapse-related outcomes, myelination, inflammatory processes, and neuroprotection in the EAE model of MS. Across EAE studies, nimodipine generally reduced clinical disease severity or cumulative burden, although relapse-related outcomes were inconsistent. CONCLUSIONS: Preclinical evidence suggests that nimodipine may attenuate disease severity and demyelination and may promote repair-related processes in rodent models relevant to MS. However, to evaluate the clinical applicability of nimodipine in MS patients, well-powered, transparently reported preclinical replication and early-phase clinical studies are required before clinical translation.

Animals

Effects of continuous isomaltulose-containing gummy intake on interstitial glucose and salivary hormones during an 18-hole golf round: a randomized, double-blind controlled pilot study.

BACKGROUND: Golf is a prolonged, moderate-intensity sport requiring sustained physiological stability to manage cumulative stress and maintain performance. Although carbohydrate intake is commonly used to reduce fatigue, rapidly absorbed sugar-induced rapid blood glucose fluctuations may induce volatile arousal and latent metabolic stress. Isomaltulose, a slow-digesting disaccharide, provides a steadier glucose supply compared with sucrose. This exploratory pilot study examined the effects of isomaltulose intake on physiological stress markers, glycemic dynamics, and subjective responses during a competitive 18-hole golf round. METHODS: Twenty-three male collegiate golfers were randomized to either the isomaltulose group (ISO; n&#x2009;=&#x2009;12) or the sucrose group (CON; n&#x2009;=&#x2009;11) in a double-blind controlled trial. Participants consumed gummies containing isomaltulose or sucrose immediately after each hole (12.1 g carbohydrate per hole; total carbohydrate intake: 217.5 g). Primary outcomes were salivary stress markers [cortisol, testosterone, and dehydroepiandrosterone sulfate (DHEAS)] levels. Secondary outcomes included interstitial glucose concentration measured via continuous glucose monitoring, subjective assessments (i.e. sleepiness, relaxation, and concentration), and golf performance (18-hole score). Between-group comparisons at each time point were conducted using planned Welch's t-tests. RESULTS: No significant between-group differences were observed for 18-hole score (p&#x2009;=&#x2009;0.38) or mean interstitial glucose concentration (p&#x2009;=&#x2009;0.20). However, exploratory analyses revealed distinct hormonal variations; salivary DHEAS and testosterone levels were higher in the ISO group during the latter half of the round (p&#x2009;<&#x2009;0.05), whereas both declined in the CON group. Regarding glycemic variability, the ISO group demonstrated a more stable glucose profile with a medium effect size for lower standard deviation (ISO: 14.7&#x2009;&#xb1;&#x2009;1.9 vs. CON: 16.7&#x2009;&#xb1;&#x2009;4.6 mg/dL; d&#x2009;=&#x2009;0.58), although this difference was not significant. Conversely, subjective outcomes diverged; the CON group reported significantly greater subjective arousal (wakefulness and relaxation) (p&#x2009;<&#x2009;0.01) relative to the ISO group. CONCLUSIONS: In conclusion, continuous intake of isomaltulose-containing gummies during an 18-hole golf round was associated with differences in selected physiological markers, including DHEAS and testosterone concentrations. However, these findings were not accompanied by improvements in objective golf performance outcomes compared with sucrose-containing gummies. Isomaltulose may influence glycemic dynamics and hormonal responses during prolonged golf play; however, the practical significance of these effects remains exploratory. Further studies with larger sample sizes and appropriate repeated-measures frameworks are needed to determine whether such physiological changes translate into meaningful performance or recovery benefits.

Humans

Evaluation of three Aspergillus antibody assays for screening of chronic pulmonary aspergillosis: prospective diagnostic accuracy study.

OBJECTIVES: Chronic pulmonary aspergillosis (CPA) is a frequent complication of pulmonary tuberculosis (PTB), particularly in high-burden settings where access to reliable serological diagnostics remains limited. We evaluated the diagnostic performance of two immunochromatographic technology (ICT) lateral flow assays (LFAs) and an ELISA for CPA screening among patients with active or previously treated PTB. METHODS: In this two-year prospective multicentre diagnostic evaluation, serum from adults with prior or active PTB was tested using the Era Biology Aspergillus IgG ICT LFA, LDBio Aspergillus IgG/IgM ICT LFA, and Bordier Aspergillus fumigatus IgG ELISA. CPA diagnosis was established using a consensus composite reference standard incorporating clinical, immunological, radiological, and microbiological criteria. The Bordier ELISA was used as part of the immunological component of the consensus CPA diagnosis, with a cutoff optical density of &#x2265;1.0. Diagnostic accuracy, agreement statistics, receiver operating characteristic analysis, and latent class analysis (LCA) were performed. RESULTS: Among 340 participants, 24 (7.06%) had CPA. Proportion of participants with positive antibody tests among all tested individuals were 6.76% for LDBio ICT LFA, 20.0% for Era Biology ICT LFA, and 11.47% for Bordier ELISA. Against consensus CPA diagnosis, Bordier ELISA showed 87.50% sensitivity and 94.30% specificity, LDBio ICT LFA 58.33% sensitivity and 97.15% specificity, and Era Biology LFA 66.67% sensitivity and 83.54% specificity. LCA estimated CPA prevalence at 7.72%. LCA-derived sensitivities and specificities were 86.58% and 99.92% for LDBio ICT LFA, 83.39% and 85.31% for Era Biology LFA, and 79.10% and 94.19% for Bordier ELISA. CONCLUSIONS: The Bordier ELISA showed high sensitivity and specificity, while the LDBio ICT LFA demonstrated very high specificity with strong LCA-derived performance. These findings support the use of ELISA for laboratory diagnosis and ICT as a point-of-care screening tool for CPA in resource-limited settings. Era Biology Aspergillus IgG LFA demonstrated moderate sensitivity and acceptable diagnostic performance, indicating its potential utility as a supplementary screening assay for CPA in settings where rapid, point-of-care testing is required.

Humans

Diagnostic value of plasma cell-free DNA metagenomic next-generation sequencing in patients with suspected infections and exploration of clinical scenarios-a retrospective study from a single center.

BACKGROUND: Plasma cell-free DNA metagenomic next-generation sequencing (mNGS) is a non-invasive comprehensive method for the etiological diagnosis of various infectious diseases. However, research on the early diagnosis and real-world clinical impact of plasma mNGS in patients with suspected infection are still limited. MATERIALS AND METHODS: This study retrospectively included 140 patients with suspected infections who underwent early plasma mNGS and conventional culture testing. Referring to the clinical diagnosis of infectious diseases, the diagnostic performance of plasma mNGS and culture tests was compared, and the application scenarios and clinical effects of plasma mNGS were evaluated. RESULTS: The positive rate of plasma mNGS was significantly higher than that of culture methods (55.71% vs 25.10%, p&#x2009;<&#x2009;0.001) and blood cultures (55.71% vs 12.86%, p&#x2009;<&#x2009;0.001). Regarding clinical diagnosis, the sensitivity of plasma mNGS was significantly higher than that of culture (58.27% vs 37.80%, p&#x2009;=&#x2009;0.002). The combination of mNGS and culture achieved a higher detection sensitivity (69.29%), especially in patients with multi-site co-infections (73.68%) and blood infections (73.17%). Plasma mNGS demonstrated higher sensitivity in patients with procalcitonin (PCT) index > 5&#x2009;ng/ml or human neutrophil lipocalin (HNL) index > 200&#x2009;ng/ml. In terms of treatment, a total of 69 patients (54.33%) benefited from plasma mNGS. CONCLUSION: This study highlights the significant improvement in pathogen detection performance by combining conventional culture with plasma mNGS detection, especially in patients with multi-site co-infections and blood infections. Early use of plasma mNGS as an adjunct to culture can better guide clinicians to initiate appropriate anti-infective therapy.

Humans

The future of pediatric vesicoureteral reflux management.

BACKGROUND AND OBJECTIVE: Vesicoureteral reflux (VUR) is a common condition in pediatric urology, yet important uncertainties persist regarding risk stratification, imaging strategies, and prevention of long-term renal damage. Emerging technologies may help address these challenges. This review provides a forward-looking overview of recent advances in artificial intelligence (AI) and immunomodulation that may influence future management of pediatric VUR. METHODS: A forward-looking literature review was performed using the PubMed database (January 2000-March 2025), focusing on studies addressing AI, immunomodulation, or vaccination in the context of VUR and urinary tract infections. Criteria of inclusion were the relevance to pediatric VUR, the novelty of the proposed concept, the potential clinical implications and, for the AI literature, the existence of a clinical evaluation of the algorithm on a dataset from patients. KEY FINDINGS AND LIMITATIONS: AI-based models show promising performance in supporting clinical decision-making, including prediction of the need for voiding cystourethrography, automated grading of VUR, estimation of recurrent urinary tract infection risk and prediction of chemoprophylaxis. These tools may facilitate more individualized diagnostic and therapeutic strategies, although current evidence is largely retrospective and requires prospective validation. Immunization and immunomodulatory approaches aim to reduce infection burden and modulate inflammatory pathways associated with renal scarring. While early experimental and adult clinical data are encouraging, pediatric-specific evidence remains limited, and clinical applicability in children with VUR is not yet established. CONCLUSION: Artificial intelligence and immunologically targeted strategies represent complementary, emerging approaches that may contribute to more personalized management of pediatric VUR. At present, both should be regarded as exploratory tools whose clinical impact will depend on further validation and appropriately designed pediatric studies.

Humans

Game changer? Cognitive-motor effects of VR exergaming compared to video-based training.

BACKGROUND/OBJECTIVE: Virtual reality (VR) exergaming enhances several cognitive domains through multisensory engagement. Acute cognitive benefits of VR are established, but evidence for direct comparisons with non-immersive controls is limited. This study aimed to determine whether VR exercise provides additional cognitive and cognitive-motor benefits beyond a matched non-immersive active stick-fight video (SFV) intervention, and whether effects persist after training. METHODS: In this randomized quasi-experimental study, N&#x2009;=&#x2009;55 healthy adults (VR: n&#x2009;=&#x2009;30; SFV: n&#x2009;=&#x2009;25; 25.5&#x2009;&#xb1;&#x2009;7.1&#x2009;years; 41.8% female) completed an 8-week program (2&#x2009;&#xd7;&#x2009;30&#x2009;min/week), of VR or SFV matched in movement patterns, frequency, intensity and duration. Measurements included reaction time (RT), Stroop Test (versions 1-3), Letter Cancellation Test (LCT), Trail Making Test (TMT), Trail Walking Test (TWT) and Fitts task (difficulty level 1-4). Data were analyzed using mixed-design ANOVAs. RESULTS: Improvements were observed in Stroop reading (F(1,53) = 14.84, p < .001, &#x3b7;2 = 0.219), Stroop inhibition (F(1,53) = 10.99, p = .002, &#x3b7;2 = 0.172), and LCT (F(1,53) = 4.57, p = .037, &#x3b7;2 = 0.079). A time&#x2009;&#xd7;&#x2009;group interaction was found for TMT (F(1,53) = 6.55, p = .031, &#x3b7;2 = 0.110), indicating greater changes following VR training. Both groups improved cognitive-motor performance (TWT: F(1,25) = 55.32, p < .001, &#x3b7;2 = 0.689; Fitts3: F(1,53) = 44.97, p < .001, &#x3b7;2 = 0.459), with greater gains for VR in Fitts3 (p = .006). CONCLUSION(S): Eight weeks of VR and SFV enhanced cognitive and cognitive-motor performance. VR provided domain-specific advantages in executive function, but these effects were not uniformly persistent. SFV sustained more improvements in real-world-relevant cognitive-motor tasks.

Humans

Prognostic value of the lactate-to-albumin ratio in adult sepsis: An updated systematic review of prognostic evidence.

BACKGROUND: The lactate-to-albumin ratio (LAR) has emerged as a potential prognostic biomarker in sepsis. This systematic review evaluated the prognostic value of LAR for mortality in adults with sepsis or septic shock. METHODS: PubMed/MEDLINE, Embase, Web of Science, Scopus, and the Cochrane Library were searched from inception through March 2026. Studies evaluating mortality-related prognostic performance of LAR in adults with sepsis or septic shock were included. Risk of bias was assessed using the Quality In Prognosis Studies (QUIPS) tool. Adjusted odds ratios (ORs) and hazard ratios (HRs) were evaluated separately because of methodological heterogeneity. Discrimination was assessed using study-specific area under the curve (AUC), sensitivity, specificity, and LAR thresholds. RESULTS: Fourteen primary studies were included. Higher LAR was consistently associated with increased mortality across emergency department and intensive care populations. AUC values generally ranged from approximately 0.65 to 0.87, although one smaller cohort reported an AUC of 0.976. Several multivariable analyses demonstrated associations between higher LAR and mortality after adjustment for clinical covariates. Adjusted ORs and HRs were not pooled because of differences in LAR scaling, thresholds, mortality endpoints, and adjustment strategies. Considerable variability was observed in reported cut-offs and diagnostic performance. CONCLUSIONS: Higher LAR is associated with mortality in adult sepsis and may provide complementary prognostic information. However, clinical and methodological heterogeneity precludes a universal cut-off or single pooled adjusted effect. Standardized prospective multicenter studies are required before routine clinical implementation.

Humans

Artificial intelligence for dental caries detection: An umbrella review.

Artificial intelligence (AI) has been proposed as a tool to improve dental caries detection across imaging modalities; however, its clinical value remains uncertain. This umbrella review aimed to synthesize and critically appraise systematic reviews evaluating AI for caries detection and diagnosis. An umbrella review was conducted following PRIOR guidance (PROSPERO CRD420261340728). Searches were performed in MEDLINE, Embase, Scopus, Web of Science, and Google Scholar up to 15 March 2026. Methodological quality was assessed using AMSTAR 2, and overlap of primary studies was quantified using the corrected covered area (CCA). Seventeen systematic reviews were included, of which five reported diagnostic test accuracy meta-analyses using bivariate or HSROC models. Across these meta-analyses, pooled sensitivity ranged from 0.76 to 0.94 and specificity from 0.85 to 0.91. Most systems were based on deep learning models applied to bitewing radiographs and intraoral photographs. However, substantial heterogeneity was observed in imaging modalities, lesion thresholds, analytical tasks, and evaluation metrics. In addition, a high degree of overlap across reviews and recurrent methodological limitations, including reliance on retrospective datasets, limited external validation, and inconsistent reporting, substantially weaken the reliability of the evidence. Although AI models demonstrate high diagnostic performance under experimental conditions, current evidence does not support their use as stand-alone diagnostic tools. Their clinical applicability remains limited, and implementation should be restricted to decision-support contexts until robust prospective validation demonstrates meaningful impact on clinical decision-making and patient outcomes.

Dental Caries

Effects of free-weight resistance training based on hexagonal barbell deadlift in older women: A 24-week randomized controlled trial.

PURPOSE: This randomized controlled trial examined effects of a 24-week hexagonal barbell deadlift (HBDL)-based free-weight resistance training program on body composition, trunk muscle function, and functional performance in older women. METHODS: Thirty-two women (67.6&#xa0;&#xb1;&#xa0;6.3&#xa0;years) were randomly assigned to an HBDL group (DG, n&#xa0;=&#xa0;16) or control group (CG, n&#xa0;=&#xa0;16). DG trained twice weekly for 24&#xa0;weeks under supervision. Primary outcomes were body composition and isokinetic trunk peak torque and average power at 60&#xb0;/s and 120&#xb0;/s. Secondary outcomes were isokinetic knee function, mobility, and maximal isotonic strength. Between-group differences after the intervention were examined by analysis of covariance (&#x3b1;&#xa0;=&#xa0;0.05). RESULTS: After intervention, DG showed greater trunk lean mass (+0.4&#xa0;kg vs. -0.2&#xa0;kg, p&#xa0;=&#xa0;0.006) than CG, and reductions in body fat percentage (-1.4% vs. +0.3%, p&#xa0;=&#xa0;0.003), total fat mass (-1.1&#xa0;kg vs. +0.3&#xa0;kg, p&#xa0;=&#xa0;0.013), and regional fat mass (trunk, gluteal, thigh; all p&#xa0;<&#xa0;0.05). Trunk extensor peak torque at 60&#xb0;/s (+26.0% vs. -6.3%, p&#xa0;<&#xa0;0.001) and average power at 60&#xb0;/s (+38.1% vs. -7.6%, p&#xa0;<&#xa0;0.001) and 120&#xb0;/s (+33.2% vs. +0.1%, p&#xa0;=&#xa0;0.002) improved significantly. DG also showed better eyes-closed static balance and 6-min walk performance (both p&#xa0;<&#xa0;0.05). Whole-body lean mass and lower-limb isokinetic strength did not differ between groups. Attendance was 89.6% with no adverse events. CONCLUSION: HBDL-based free-weight resistance training is a safe, feasible, and effective strategy to improve body composition, trunk extensor function, and mobility in older women.

Humans

Temporal redistribution of control reveals age-related differences in task switching at the level of preparation.

Task-switching studies often report minimal age-related differences in switch costs, leading to the conclusion that switching-related control processes are relatively preserved in aging. However, this conclusion is based on paradigms that confound preparatory and execution processes. This study examined whether age-related differences in semantic task-set reconfiguration may be underestimated due to this confound. In Experiment 1 (36 young and 30 older adults), participants performed an externally paced task-switching paradigm without control over preparation. In Experiment 2 (28 young and 28 older adults), a self-paced paradigm allowed participants to initiate stimulus onset, enabling measurement of preparation time. Across both experiments, reaction time (RT) and error rate (ER) showed reliable age effects but no interactions between age and condition, whereas switching-related condition effects varied across measures and experiments. The expression of switching-related costs differed across measures and task structures. Local switch costs were expressed in ER in Experiment 1 but in RT in Experiment 2. Global switch costs (all-switch vs. all-repeat) were observed in execution measures only in Experiment 1. In Experiment 2, preparation time showed reliable mixing, local, and global switching effects, with age-related amplification emerging specifically for global switching. These findings indicate that switching-related costs are redistributed across processing stages and behavioral measures. The results suggest that age-related modulation of semantic task-set reconfiguration may emerge more clearly during preparation than task execution, particularly under continuous switching demands. Preparation time is interpreted cautiously as reflecting participant-regulated preparatory processes rather than a pure measure of preparation efficiency.

Humans

Construction of precision clinical-proteomics risk model based on machine learning for predicting heart failure in type II diabetes mellitus.

BACKGROUND AND AIMS: Heart failure (HF) is a severe complication in type 2 diabetes mellitus (T2DM), but current risk stratification scores have limited predictive accuracy. We aimed to develop novel prediction tools integrating clinical variables with proteomics to improve risk stratification of hospitalization for HF in T2DM. METHODS AND RESULTS: In this study, we included 2111 UK Biobank participants with T2DM but no prior HF, and profiled 2920 proteins to predict 10-year incident HF hospitalization. Participants were randomly divided into training (70%), tuning (10%), and validation (20%) sets.Three prediction models were developed: a Clinical model based on demographic characteristics, comorbidities, medication use, and laboratory indices; a Protein model based on 40 proteins selected by the Light Gradient Boosting Machine (LGBM); and the Clinical OMics and Protein ASSessment for Heart Failure (COMPASS-HF) model, which integrated both clinical variables and the LGBM-selected proteins. Models were evaluated for area under the curve (AUC), sensitivity, and specificity. During follow-up, 168 participants (7.96%) developed incident HF. The COMPASS-HF model showed better discrimination than the Clinical model, with an AUC of 0.897 (95% CI: 0.850-0.945) versus 0.790 (95% CI: 0.723-0.856). It also demonstrated higher sensitivity (0.882; 95% CI: 0.725-0.967) and consistent performance in subgroups. COMPASS-HF effectively stratified risk of hospitalization for HF, with cumulative incidence rates of 31.9% in the high-risk group and 1.2% in the low-risk group. CONCLUSIONS: By combining clinical and proteomic variables, we developed a high-performance HF prediction model for T2DM, enabling precise risk stratification and informing early intervention strategies.

Humans

Artificial Intelligence for Diagnosis, Risk Stratification, and Prognosis of Neuroblastoma - A Systematic Review and Meta-Analysis.

PURPOSE: To synthesizes evidence on artificial intelligence (AI) performance in neuroblastoma (NB) diagnosis, risk stratification, prognosis, and genomic characterization. MATERIALS AND METHODS: A systematic review and meta-analysis was conducted following PRISMA 2020 guidelines (PROSPERO: CRD42024539475) across five databases. Meta-analyses used random-effects models with logit-transformed Area Under the Curve (AUCs) and cluster-robust standard errors. AI models were classified as Machine Learning Models (MLM) or Hybrid Nomograms (HN) based on their construction methodology. RESULTS: Of 3,742 articles identified, 53 were included. MLMs demonstrated higher point estimates than radiologists in differential diagnosis (AUC: 0.87 vs. 0.83), though this difference was not statistically significant and carried substantial uncertainty. HNs achieved stronger performance in risk stratification (AUC: 0.87). AI-derived nomograms (AUC: 0.9) and gene signatures (AUC: 0.8) outperformed conventional prognostic markers descriptively. Chemotherapy response prediction remained below clinical utility thresholds across all model types. Only 33.9% of models reported calibration and 24.5% underwent external validation. CONCLUSIONS: AI demonstrates proof-of-concept across multiple NB clinical domains. However, clinical adoption remains premature given persistent gaps in external validation, calibration, dataset size, and pediatric-specific model development. Future studies should test these models prospectively in multicenter pediatric cohorts, ideally through COG or SIOPEN, using shared definitions for diagnosis, risk group, treatment response, and survival outcomes.

Humans

Structured robotic colorectal training in a non-tertiary NHS hospital: a 502-case consecutive cohort implementation study.

Robotic-assisted colorectal surgery has expanded rapidly across NHS practice in the UK. Structured unit-wide training pathways are essential for safe technology adoption, yet published outcome data from non-tertiary hospitals remain limited. This study describes the implementation and feasibility of a unit-wide robotic colorectal program at a high-volume non-tertiary hospital, reporting outcomes across 502 consecutive resections performed by eight consultant surgeons and presenting these in the context of nationally published benchmarks. A retrospective cohort study of 502 consecutive robotic colorectal resections performed at York Teaching Hospital between May 2022 and December 2025. Eight consultant surgeons (A-H) participated in a structured four-phase training pathway incorporating simulation training, proctored cases, complexity-based case progression, and formal credentialing. Primary outcomes were 30-day mortality, unplanned return to theatre (RTT), and anastomotic leak (AL). Anastomotic leak was calculated using only patients who underwent anastomosis as the denominator. Procedure-stratified and individual surgeon outcomes with 95% confidence intervals were reported. Risk-adjusted cumulative sum (RA-CUSUM) analysis was performed to evaluate learning curves. Outcomes are presented descriptively alongside nationally published reference data; no formal statistical comparison against national benchmarks was performed. 502 robotic colorectal resections were performed. Mean patient age was 70.0 &#xb1; 11.3&#xa0;years; 58.4% were male. Median ASA grade was III. The indication was malignancy in 89.2% of cases. Length of stay was non-normally distributed and is therefore reported using median and interquartile range in the revised analysis. Key outcomes: - 30-day mortality: 1.0% (5/502; 95% CI 0.4-2.3%) - Unplanned return to theatre (RTT): 5.2% (26/502; 95% CI 3.6-7.5%) - Anastomotic leak (AL): 3.3% (15/450; 95% CI 2.0-5.5%; denominator = patients with anastomosis) - 30-day unplanned readmission: 5.0% (25/502; 95% CI 3.4-7.2%) - Conversion to open surgery: 3.6% (18/502; 95% CI 2.3-5.6%) - Lymph node yield &#x2265;12: 91.3% of cancer resections - R0 resection rate: 95.1% of cancer resections All primary outcomes fell within or below the published reference ranges used for descriptive context. RA-CUSUM trajectories were heterogeneous: no surgeon crossed the predefined upper control limit, but several curves showed later upward movement. Accordingly, the analysis is interpreted as safety surveillance rather than evidence of uniform performance improvement. RA-CUSUM monitoring showed that no surgeon crossed the predefined upper control limit; however, heterogeneous trajectories precluded a claim of uniform performance improvement.

Humans

Heterogeneity in Teriflunomide Treatment Arms: A Systematic Review and Meta&#x2011;Regression of Randomised Multiple Sclerosis Trials.

BACKGROUND: Teriflunomide is widely used as an active comparator in Phase 3 randomised trials for relapsing multiple sclerosis (RMS). Temporal changes in disease activity within teriflunomide-treated cohorts have not been systematically examined. OBJECTIVES: To assess temporal trends in relapse and disability outcomes across teriflunomide arms of Phase 3 multiple sclerosis (MS) trials and identify predictors of between-trial heterogeneity. METHODS: We performed a systematic review and meta-analysis of Phase 3 randomised controlled trials including a teriflunomide arm. PubMed, Scopus, and ClinicalTrials.gov were searched up to October 2025. Annualised relapse rate (ARR) and 12- and 24-week confirmed disability worsening (CDW) were extracted together with baseline characteristics. Risk of bias was assessed using the Cochrane Risk of Bias 2 tool. Random-effects meta-analyses, meta-regression, and sensitivity analyses were performed. RESULTS: Twelve teriflunomide cohorts from eight trials involving 4,900 adults with RMS were included. ARR ranged from 0.11 to 0.37 with substantial heterogeneity (I2 = 94%). Trial start year was inversely associated with ARR and explained a large proportion of between-study variability in exploratory meta-regression analyses. Confirmed disability worsening outcomes also showed substantial heterogeneity with a weaker trend toward lower event rates in more recent trials. CONCLUSION: Teriflunomide-treated trial populations have shifted toward lower relapse activity over time, and trial start year was the principal predictor of between-trial heterogeneity in ARR in exploratory analyses. These findings most plausibly reflect evolving recruitment and diagnostic practices rather than changes in drug efficacy. Accounting for these temporal dynamics is essential when interpreting outcomes from RMS trial using teriflunomide as comparator.

Humans

Fused Deposition Modeling (FDM) of polyether-ether-ketone (PEEK) dental implants: A systematic review of the effect of printing parameters on mechanical behaviour and surface quality.

PURPOSE: This systematic review evaluated how FDM printing parameters influence mechanical behaviour and surface characteristics of 3D-printed PEEK and identified parameter combinations linked to the most favourable mechanical performance and surface quality. MATERIALS AND METHODS: An electronic search was conducted in: MEDLINE (Ovid), PubMed, Embase, Web of Science, Scopus, and Compendex (last update: January 2025). Studies that evaluated the effect of FDM printing parameters on mechanical and surface properties of PEEK were included. Outcomes comprised compressive, tensile, and flexural strengths, elastic modulus, fracture toughness, surface hardness, roughness, and wettability. RESULTS: Of 4005 reports screened, 54 manuscripts were included. 92.6% (n&#x202f;=&#x202f;50) of articles showed low risk-of-bias, while 7.4% (n&#x202f;=&#x202f;4) showed medium risk-of-bias. Tensile strength was the most investigated mechanical parameter (78%), followed by elastic modulus (41%), flexural strength (30%), compressive strength (20%), and fracture toughness (6%). Surface roughness was the most evaluated surface property (30%), followed by hardness (17%) and wettability (6%). Across studies, higher printing temperatures, lower printing speed, thinner layer thickness, and maximum infill ratio in a horizontal printing orientation were associated with higher strengths, less warpage, increased accuracy, and improved surface quality. CONCLUSION: Specific combinations of FDM printing parameters can significantly improve the mechanical and surface properties of PEEK. However, it is difficult to meet all the optimal conditions simultaneously. Thus, balancing between different parameters must be considered in practical production.

Benzophenones

Plasma proteomics reveal SERPINA1 and CD59 as candidate biomarkers for COVID-19 severity stratification and prognosis prediction.

BACKGROUND: COVID-19 has been closely associated with coagulation abnormalities. However, existing biomarkers, including D-dimer and fibrin degradation products (FDP), exhibit limited accuracy in stratifying disease severity and predicting long-term clinical outcomes. OBJECTIVES: This study aimed to use proteomic analysis to identify plasma biomarkers associated with COVID-19 severity and prognosis, and validate their predictive utility for mortality and thromboembolic complications. METHODS: Plasma proteomic profiles were analyzed across three COVID-19 severity classes. Differential expression analysis and functional analysis were performed. Clustering analysis was used to identify proteins correlated with disease severity. Candidate biomarkers were validated in an independent cohort. Predictive performance of the biomarkers for mortality, sepsis and venous thromboembolism was evaluated using bootstrap-corrected ROC analyses and multivariable regression analyses. RESULTS: Proteomic analysis revealed progressive involvement of the coagulation and complement pathway with increasing disease severity. SERPINA1 and CD59 were identified as candidate biomarkers and exhibited significantly higher plasma levels in severe cases. Bootstrap-corrected ROC analyses demonstrated strong predictive performance: SERPINA1 achieved AUCs of 0.775 and 0.924 for 30-day and 12-month mortality, and CD59 achieved AUCs of 0.720 for sepsis; the combined model further improved prediction of 12-month mortality (AUC 0.946) and sepsis (AUC 0.904), outperforming D-dimer and FDP. Multivariable regression confirmed their independent prognostic value. CONCLUSION: This exploratory study identifies SERPINA1 and CD59 as candidate prognostic biomarkers in COVID-19, highlighting the role of coagulation and complement-related pathways in disease severity and warranting further prospective validation.

Humans