Search PubMedSearch

SEARCH · Search PubMed

Results for “interreviewer agreement”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

40 records · Page 2Linked to original sources

Efficacy of the NMIC-150 system in identifying extended-spectrum beta-lactamases in clinical isolates.

Extended-spectrum beta-lactamases (ESBLs) are significant contributors to the growing global crisis of antimicrobial resistance. This study evaluated the performance of the NMIC-150 System for susceptibility testing of third-generation cephalosporins (3GCs) and assessed whether ceftazidime-avibactam and aztreonam-avibactam could identify ESBL-producing carbapenem-resistant Enterobacterales (CREs). A total of 278 non-duplicate clinical isolates (Klebsiella pneumoniae, E. coli, and Proteus mirabilis) were analyzed. Antimicrobial susceptibility was determined using reference broth microdilution (BMD) and the NMIC-150 System. ESBL production was defined as an ≥eight-fold reduction in the minimum inhibitory concentration (MIC) of 3GCs in the presence of clavulanic acid, according to CLSI criteria. Whole-genome sequencing was performed to characterize ESBL and carbapenemase genes among 3GC-resistant isolates. A Random Forest model was used to predict ESBL-producing isolates based on MIC values. The NMIC-150 System demonstrated over 90% categorical and essential agreement with BMD for ceftazidime and ceftriaxone, along with robust predictive performance via Random Forest analysis. These findings suggest that the NMIC-150 System is a reliable platform for 3GC susceptibility testing and that an ≥eight-fold MIC reduction with ceftazidime-avibactam or aztreonam-avibactam may serve as a phenotypic indicator of ESBL production in CRE isolates. In conclusion, the NMIC-150 System shows potential for routine antimicrobial resistance surveillance and may facilitate the rapid identification of ESBL-producing CREs in clinical settings.

Microbial Sensitivity Tests

An AI-assisted Clinical Decision Support System for Green Classification of Cystocele on Dynamic Transperineal Ultrasound.

Green classification of cystocele on dynamic transperineal ultrasound (TPUS) remains operator-dependent because it requires manual frame selection and landmark-based assessment of the Valsalva maneuver. We developed a workflow-oriented AI-assisted clinical decision support system for automated urethrovesical junction localization and dynamic Green classification and prospectively evaluated its standalone and reader-support performance. This diagnostic accuracy and reader study included 881 patients from a tertiary referral hospital, comprising a retrospective development cohort (n = 688) and an independent prospective test cohort (n = 193). A nested subset of 67 prospective patients was used for a reader study involving two junior and two intermediate radiologists under unaided and AI-assisted conditions. In the complete prospective test cohort, Green-AttGRU achieved a macro-averaged AUC of 0.939 (95% CI, 0.897-0.971) and an overall accuracy of 0.902 (95% CI, 0.860-0.943). In the reader study, overall accuracy increased from 0.761 to 0.821 without AI to 0.851-0.881 with AI, while macro-F1 increased from 0.660 to 0.777 to 0.820-0.860. Overall inter-reader agreement increased from a Fleiss' κ of 0.453 to 0.786, and pooled median interpretation time decreased from 26.7 s to 9.9 s. These findings support the preliminary feasibility of the system as a workflow-oriented decision-support tool for dynamic TPUS interpretation.

Humans

Instruments to Assess Bidirectional Intimate Partner Violence: A Systematic Review.

Intimate partner violence (IPV) is a significant public health issue that affects individuals, families, and society in profound ways. Over the years, research has often shown that such IPV is predominantly one-sided, with men as perpetrators and women as victims. However, more recent studies have revealed that IPV is frequently bidirectional, with both partners potentially being victims, perpetrators, or both. Despite this growing awareness, little is known about the tools used to assess bidirectional violence (BV). This systematic review aims to synthesize the published scientific literature on the instruments used to assess bidirectional IPV among adult men and women. Following Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) guidelines, a search was conducted across databases such as PubMed, B-on, Web of Science, and Scielo. Fifty studies published between 2012 and 2024 met the inclusion criteria. Most of the studies were conducted with community samples recruited through online questionnaires. The Conflict Tactics Scale emerged as the most used instrument for evaluating BV, demonstrating robust psychometric properties. BV prevalence rates ranged from 2% to 98.4% and remained generally consistent across different instruments, except in the case of male reports. However, the limited number of studies on interpartner agreements and the reliance on single-partner self-reports pose challenges to the accuracy of these estimates. While this systematic review provides a comprehensive overview of the instruments used to assess BV, it underscores the need for future research to develop more precise, context-sensitive tools, that incorporate reports from both partners to improve the accuracy of BV assessment.

Humans

Psychometric Evaluation of the Breast Inflammatory Symptom Severity Index Versions 2 and 3 Among Lactating Women.

OBJECTIVE: To evaluate the psychometric properties of Versions 2 and 3 of the Breast Inflammatory Symptom Severity Index (BISSI). DESIGN: Secondary data analysis of clinical trial data. SETTING: Private physiotherapy practices, a public tertiary hospital, and a community in Melbourne, Australia. PARTICIPANTS: Women more than 7 days after birth with inflammatory conditions of the lactating breast (N = 43). METHODS: We performed confirmatory factor analysis of the BISSI Version 2 to examine item loading, which informed development of the BISSI Version 3 (V3). We assessed convergent validity by comparing total BISSI V3 scores with human milk sodium to potassium ratio (Na+:K+) at Trial Days 1, 3, and 10 using Bland-Altman plots. We compared item-level scores for size of affected area with objective receiver operating characteristic curve analysis to assess discriminant validity for symptom severity and Cronbach's alpha for internal reliability. RESULTS: After confirmatory factor analysis, we removed two items, resulting in a six-item BISSI V3. All retained items demonstrated comparable loading on the overall scale. Limits of agreement for total BISSI V3 scores and item-level scores for size of affected area were acceptable at all time points, with more than 90% of observations falling within 2 standard deviations of the mean difference, supporting convergent validity. Discriminant validity of the BISSI V3 was supported. We found high internal reliability at both time points CONCLUSION: Our findings provide evidence for the validity and reliability of the BISSI V3 and support its continued development and for clinical use of the BISSI V3 and human milk Na+:K+ analysis to enhance management of inflammatory conditions of the lactating breast.

breastfeeding

An individualized nomogram for predicting progression-free survival in systemic anaplastic large cell lymphoma: a multicenter, retrospective, and internally validated study.

OBJECTIVES: To develop an individualized nomogram for predicting disease progression risk in systemic anaplastic large cell lymphoma (sALCL). METHODS: Independent predictors of progression-free survival (PFS) were identified using Cox regression in a multicenter retrospective cohort of 109 sALCL patients (2010-2022). These were incorporated into a three-factor nomogram, evaluated via bootstrapped internal validation (1000 resamples), ROC analysis, C-index, decision curve analysis (DCA), and clinical impact curve (CIC). RESULTS: A total of 29 PFS events occurred during a median follow-up of 31 months. Multivariable modelling selected serum β2-microglobulin elevation, extranodal disease, and front-line chemotherapy choice (CHOP versus CHOPE or BV+CHP) as autonomous progression drivers. Upon internal bootstrap validation, the nomogram yielded strong prognostic accuracy, achieving AUCs of 0.81, 0.85 and 0.87 for 1-, 3- and 5-year progression-free survival, alongside a corrected C-index of 0.779 (95% CI: 0.699 - 0.861). Calibration plots showed close agreement between predicted and observed outcomes, while DCA confirmed superior net clinical benefit versus conventional IPI or Ann Arbor stratification across multiple decision thresholds. CONCLUSION: This first sALCL-specific nomogram integrates clinical and treatment variables to provide personalized PFS risk estimation. While internally validated, this exploratory, observation-based tool requires external validation and recalibration in prospective cohorts before clinical implementation.

Humans

Evaluation of the difference between automated and measured QTc intervals in children.

BACKGROUND: The corrected QT interval (QTc) is obtained through automated ECG computations or manual physician measurements. We hypothesized that differences exist in children between the measured and automated QTc intervals within and between Healthy and hypertrophic cardiomyopathy (HCM) subjects with greater differences for HCM due to structural abnormalities. METHODS: QT measurements - Bazett correction- automated (aQTc) and measured (mQTc), were extracted from the GE MUSE database for 385 Healthy pediatric (single ECG) and 208 HCM subjects (2 ECGs), stratified by age&#xa0;<&#xa0;12 and&#xa0;&#x2265;&#xa0;12&#xa0;yrs., sex, race, and ethnicity. QTc means (SD), automated and measured differences, and the difference of the differences of aQTc and mQTc were analyzed overall and by subgroups. All ECGs were read by one pediatric cardiologist with a second cardiologist reading a random subset of HCM ECGs to evaluate intraclass correlations and agreement. RESULTS: The mQTc intervals were shorter than aQTc intervals within Healthy (p&#xa0;<&#xa0;0.001) and within first HCM ECGs (p&#xa0;<&#xa0;0.001) with both aQTc and mQTc shorter in Healthy than HCM (p&#xa0;<&#xa0;0.001). The difference in these differences was significant overall using HCM ECG 1 but not HCM ECG 2. Healthy subject aQTc and mQTc intervals differed by age, sex, and race (p&#xa0;<&#xa0;0.002). HCM ECG 1 aQTc- mQTc intervals differed for age&#xa0;<&#xa0;12&#xa0;yrs., as well as by sex and race. HCM ECG 2 intervals differed only for age&#xa0;<&#xa0;12&#xa0;yrs. CONCLUSIONS: Compared to measured values, automated QTc values were significantly longer in both Healthy and HCM subjects. Automated measurements may overestimate the QTc.

Humans

Microbial signal profiles and organism-level concordance between plasma metagenomic sequencing and blood culture in suspected bloodstream infection.

Plasma metagenomic next-generation sequencing (mNGS) and blood culture detect different components of the microbial signal and frequently produce discordant organism reports. We characterized microbial signal class, report-derived burden, organism-level concordance, and independent clinical attribution in a retrospective, single-center, episode-level cohort. Among 329 episodes with evaluable plasma mNGS reports, 315 had blood culture performed; 232 were mNGS positive/culture negative and 53 were positive by both methods. In the 232 discordant episodes, the recorded routine-care diagnosis classified 124 as bloodstream infection (BSI) and 108 as non-BSI. Nonviral signals were present in 78.2% and 42.6%, respectively (P&#x2009;<&#x2009;0.001), and median maximum report-derived sequence counts were 98.5 and 11.5 (P&#x2009;<&#x2009;0.001). Two laboratory physicians then independently reviewed source records using structured criteria while masked to the recorded BSI label and mNGS organism and sequence-count information. Initial agreement for the five-category BSI assessment was 97.6% (Cohen's kappa, 0.960). Within the mNGS-positive/culture-negative subgroup, adjudicated BSI likelihood showed a modest ordinal association with report burden (Spearman rho&#x2009;=&#x2009;0.190; P&#x2009;=&#x2009;0.004), while mNGS organisms were considered supported in 1 episode, plausible in 158, unlikely or contaminant in 72, and unresolved in 1. Among 53 dual-positive episodes, 33 (62.3%) shared at least one species, but only 5 (9.4%) had complete species-set concordance. Plasma mNGS and blood culture therefore frequently generated non-equivalent organism sets. Signal class and report burden contributed graded contextual evidence, but organism-level attribution required clinical review and orthogonal microbiology rather than binary positivity alone.

Humans

Assessment of atypical glandular cell interpretation in Pap tests using the Hologic Genius Digital Diagnostics System.

Atypical glandular cells (AGC) are a diagnostic challenge. The aim of this study was to evaluate the efficacy and diagnostic performance of AGC detection on the Hologic Genius Digital Diagnostics System (HGDDS). A retrospective analysis of 451 ThinPrep Pap cases was conducted, including 207 cases of AGC, 27 cases of high-grade squamous intraepithelial lesion (HSIL), 25 cases of low-grade squamous intraepithelial lesion (LSIL), and 192 benign cases. All AGC cases had follow-up histologic diagnoses, with 66 cases subsequently diagnosed as adenocarcinoma. The slides were randomized, scanned, and analyzed by the HGDDS. Patient age and HPV test results were provided to reviewers, an experienced cytologist, who screened the cases, followed by two cytopathologists who independently examined the cases on the HGDDS. Diagnostic concordance between the two cytopathologists indicated strong agreement (&#x3ba; = 0.829). Sensitivity of AGC on Papanicolaou (Pap) tests for adenocarcinoma detection on HGDDS was 98.5% and 95.5%, respectively, comparable to the original ThinPrep interpretation (OTPI). Specificity for adenocarcinoma detection was significantly higher (84.6% and 85.6%) with the HGDDS than 27.7% with OTPI. Overall, the diagnostic performance for AGC/HSIL interpretation to detect CIN2/3/adenocarcinoma appeared to have improved with HGDDS compared with OTPI, particularly for specificity and positive predictive value (PPV). This is the first study evaluating AGC diagnosis using the HGDDS. The findings demonstrate that the sensitivity of adenocarcinoma detection as AGC on HGDDS is comparable to the ThinPrep Imaging System, but the specificity and PPV are improved. This suggests the potential of artificial intelligence to augment the performance of cervical cancer screening.

Humans

Anabolic androgen therapy in critically ill adults: A systematic review and meta-analysis.

Critical illness is characterized by a catabolic, proinflammatory state. Anabolic agents, such as testosterone, have therefore been proposed as therapeutic targets. Our objectives were to assess the effects of testosterone in critically ill populations on patient-important outcomes and identify design limitations to inform future studies. We searched for randomized control trials (RCTs) through Medline, Embase, and EBM Reviews databases from inception through February 24, 2026, including English language articles enrolling adults (&#x2265;18&#xa0;years) admitted to ICU where anabolic androgen therapies (AAT) were compared with placebo or standard of care. Studies had to report at least one of: mortality, ICU and hospital lengths of stay, or duration of mechanical ventilation. We extracted data independently using a standardized data extraction tool, and feedback was received from all co-authors to ensure agreement. For each outcome, we performed meta-analyses using a random-effects model with inverse variance weighting in RevMan. We used the GRADE approach to assess certainty in pooled estimates of effect. Of 1325 screened articles, we found 4 that fit our inclusion criteria. Together, we judged risk of bias as 'some concerns' in 3 trials and 'high' in the final trial, and ultimately found that the effects of anabolic-androgen therapy on patient-important outcomes uncertain. With the uncertainty of current evidence for the effects of anabolic-androgen therapy in critically ill adults, there is insufficient support for its routine use. Future randomized evidence is needed to determine whether anabolic-androgen therapy improves clinically-important outcomes and better define its safety profile in critically ill adults.

Humans

Near and distance vergence facility provides complementary clinical information in concussion-related convergence insufficiency.

PURPOSE: To compare near and distance vergence facility testing in adolescents and young adults with concussion-related convergence insufficiency and evaluate changes following office-based vergence/accommodative therapy (OBVAM). METHODS: This secondary analysis of the CONCUSS randomized clinical trial evaluated vergence facility at near (40&#xa0;cm) and distance (4&#xa0;m) using a 12&#x394; base out/3&#x394; base in prism flipper. Participants aged 11-25&#xa0;years with concussion-related convergence insufficiency were randomized to immediate or 6&#xa0;weeks delayed OBVAM. Vergence facility was assessed at baseline, outcome time 1 assessment (after 12 therapy sessions for the immediate group and 6&#xa0;weeks of watchful waiting for the delayed group), and outcome time 2 assessment (after both groups completed 16 therapy sessions). Agreement between near and distance vergence facility classifications was evaluated, and treatment-related changes were compared between groups. RESULTS: Of the 106 enrolled participants, 102 completed all study visits. At baseline, the near and distance vergence facility classifications demonstrated substantial discordance. Among 101 participants with both measures available, 49 demonstrated reduced distance vergence facility despite normal near vergence facility, whereas only two showed the opposite pattern (Cohen's &#x3ba;&#xa0;=&#xa0;0.12; p&#xa0;<&#xa0;0.0001). Vergence facility improved following therapy in both treatment groups, with larger early improvements in the immediate-treatment group. CONCLUSIONS: Near and distance vergence facility testing provided complementary rather than interchangeable clinical information in adolescents and young adults with concussion-related convergence insufficiency. Both measures improved following vergence/accommodative therapy, supporting consideration of both testing distances in clinical assessment.

Humans

Can ChatGPT Replace Human Clinical Coders? A Comparative Study in Otology Billing.

OBJECTIVE: Evaluate the utility of the large language model (LLM), ChatGPT, for the analysis of operative notes and the generation of Current Procedural Terminology (CPT) codes in comparison to human clinical coders. STUDY DESIGN: CPT billing codes assigned by ChatGPT were compared to existing billing data. Otology practice within a tertiary academic center. METHODS: About 191 operative notes from a single surgeon (9/2022-10/2023) were analyzed. ChatGPT-3.5 and 4 models were prompted for CPT codes based on operative notes. Assessment included determining exact and partial match rates, sensitivity and specificity for targeted procedures, and work Relative Value Units (wRVU) differences between ChatGPT-generated and human-assigned codes. RESULTS: ChatGPT-3.5 achieved exact matches in 22% of cases and partial matches in 32%, while ChatGPT-4 achieved 14% exact and 33% partial matches. When cochlear implantation (CI) was excluded, performance dropped significantly. For CI, ChatGPT-3.5 demonstrated a sensitivity of 94% and specificity of 90%, while ChatGPT-4 showed a sensitivity of 96% and specificity of 92%. In contrast, performance on cartilage grafting was poor, with sensitivities of 4.2% for ChatGPT-3.5 and 0% for ChatGPT-4. ChatGPT-3.5 and 4 showed moderate CPT code matching accuracy among themselves, with slight agreement to human coders. Both models tended to underbill for wRVUs compared to human coders, with significant differences in the values generated. CONCLUSION: This study assessed ChatGPT's effectiveness in automating CPT code assignment for otologic surgeries. While the models achieved high sensitivity values for assigning codes related to cochlear implantation, both models struggled with complex cases, failed to apply modifiers, and often assigned fewer wRVUs. The findings highlight ChatGPT's potential in medical billing but indicate a need for further refinement.

Humans

Spatial proximity or vector orientation? Re-evaluating ECG interpretation in anterior myocardial infarction using cardiac magnetic resonance.

BACKGROUND: The electrocardiogram (ECG) is widely used to infer infarct location and extent in anterior myocardial infarction (MI), based on either anatomical lead proximity or vectorial orientation of ST-segment deviation. However, the validity of these approaches against direct imaging of myocardial injury remains uncertain. METHODS: In this prospective study, 105 patients with anterior MI underwent cardiac magnetic resonance (CMR) imaging 3-7&#xa0;days after presentation. Admission ECGs were analyzed using (1) conventional ECG localization categories, and (2) simplified frontal and horizontal ST-axis orientation. CMR-defined injury distribution was assessed using late gadolinium enhancement and myocardial edema imaging. RESULTS: Conventional ECG localization categories demonstrated no significant association with CMR-defined infarct distribution (P&#xa0;=&#xa0;0.24), with poor agreement (&#x3ba;&#xa0;=&#xa0;0.122) and substantial overlap across categories. Simplified ST-axis orientation showed modest and inconsistent associations with infarct location and did not meaningfully explain infarct size. In contrast, global ST-segment burden was associated with CMR-defined infarct size (&#x3a3;STE: standardized &#x3b2;&#xa0;=&#xa0;0.307, P&#xa0;=&#xa0;0.002; lead count: standardized &#x3b2;&#xa0;=&#xa0;0.267, P&#xa0;=&#xa0;0.007). CONCLUSIONS: In this selected cohort of reperfused LAD-related anterior STEMI patients undergoing early CMR, conventional ECG localization categories and simplified ST-axis orientation showed poor or inconsistent correspondence with CMR-defined infarct distribution, whereas global ST-segment burden showed a modest association with infarct size. These findings suggest that, in this cohort, the ECG may be better suited to reflect the extent of myocardial injury rather than its precise anatomical location.

Humans

Clinical performance of two lithium disilicate CAD/CAM materials in posterior Class II inlay restorations: A 48-month randomised split-mouth clinical trial.

OBJECTIVES: To compare the clinical performance of Amber Mill (AM) and IPS e.max CAD (EM) lithium disilicate computer-aided design/computer-aided manufacturing (CAD/CAM) materials in posterior Class II inlay restorations and characterise their baseline properties. METHODS: Thirty-four adults received paired AM and EM posterior Class II inlays (68 restorations) in a triple-blind randomised split-mouth trial followed for 48 months. Restorations were evaluated at baseline and annually using revised World Dental Federation (FDI) criteria, with fracture and retention as the primary endpoint. Baseline characterisation included flexural strength, shear bond strength, translucency parameter, and scanning electron microscopy. McNemar, Wilcoxon signed-rank, Friedman, one-way analysis of variance, Tukey post hoc, and inter-rater agreement analyses were used. RESULTS: At 48 months, 18 paired participants were available for primary analysis. Failures occurred in 2 of 18 AM restorations and in 3 of 18 EM restorations, corresponding to success rates of 88.9% and 83.3%, respectively, with no significant between-material difference (McNemar p = 1.000). No catastrophic bulk ceramic fracture was observed. Secondary FDI scores remained mostly within the clinically acceptable range; marginal staining deteriorated over time in both groups (p < .001) without significant between-material differences. Baseline material testing showed significant material- and translucency-dependent differences in flexural strength, shear bond strength, and translucency. CONCLUSIONS: Within the limitations of the 48-month follow-up and the tested Class II inlay indication, AM showed clinical performance comparable to EM. Observed clinical complications were related to retention or marginal/interface behaviour. CLINICAL SIGNIFICANCE: For posterior Class II lithium disilicate CAD/CAM inlays, medium-term complications were mainly retention/interface-related, suggesting adhesive-interface durability may be as important as baseline ceramic strength.

Humans

The musculoskeletal pain literacy questionnaire (MSK-PLq) - Part 1: Development of a preliminary version through a systematic review and Delphi consensus.

OBJECTIVE: Chronic musculoskeletal (MSK) pain is a leading cause of disability worldwide, and self-management is a first-line approach recommended by international clinical guidelines. Access to evidence-based information that enhances health literacy may support patients' engagement in their self-management and treatment decision-making, potentially reducing disease burden and pain. However, no tool currently exists to assess health literacy specifically in MSK pain. This study aimed to develop and describe the preliminary version of a knowledge-based questionnaire to evaluate MSK pain literacy, the Musculoskeletal Pain-Literacy questionnaire (MSK-PLq). METHODS: A systematic literature review identified existing health literacy instruments and generated a preliminary list of domains. A two-round Delphi study with 22 panellists (19 experts and three people living with chronic MSK pain), followed by consensus meetings, was used to refine domains and items (&#x2265;70% agreement). Readability was assessed using the Flesch Reading Ease (FRE) score and three stakeholders were consulted to review the questionnaire for comprehensibility, clarity, and face validity. RESULTS: Six domains were retained (Understand, Access, Appraise, Apply, Digital, Beliefs), comprising 20 items in the preliminary version of MSK-PLq. Readability was acceptable (mean FRE 74, indicating fairly easy reading), and subject feedback supported the questionnaire's clarity and face validity. CONCLUSIONS: The preliminary version of the MSK-PLq is proposed as the first knowledge-based tool to assess functional, interactive, and critical aspects of MSK pain literacy. It may have applications in clinical practice, research, education, and digital health, by informing tailored patient education and supporting self-management strategies, although further psychometric validation is required.

Humans

A Dynamic Nomogram to Predict Metabolic Dysfunction-Associated Fatty Liver Disease in Patients with Metabolic Syndrome.

BACKGROUND: Metabolic syndrome (MetS) involves multiple metabolic disorders. This study aimed to identify high-risk populations for metabolic dysfunction-associated fatty liver disease (MAFLD) in patients with MetS and to establish a dynamic predictive nomogram. METHODS: A total of 627 patients with MetS from six regions in Zhejiang Province were enrolled and categorized into MAFLD and non-MAFLD groups, then randomly assigned to training and validation sets at a ratio of 7:3. Independent predictors of MAFLD were identified using least absolute shrinkage and selection operator regression and multivariable logistic regression analyses. These predictors were then used to construct a dynamic nomogram. RESULTS: A total of 627 patients with MetS were included in the final analysis, of whom 77.0% (483/627) were diagnosed with MAFLD. Multivariable logistic regression analysis identified body mass index (BMI), waist circumference (WC), total cholesterol (TC), alanine aminotransferase (ALT), MetS-defined dysglycemia, and education level as independent risk factors for MAFLD. MetS-defined dysglycemia showed the highest odds ratio (OR) for MAFLD development [OR = 1.87, 95% confidence interval (CI): 1.07-3.29]. Although the number of MetS components and the metabolic syndrome score were significantly associated with MAFLD in univariate analysis, they were not independently associated with MAFLD in the multivariate model. A dynamic nomogram for predicting MAFLD risk in patients with MetS was developed and internally validated. The area under the receiver operating characteristic curve was 0.834 (95% CI: 0.787-0.880) in the training set and 0.839 (95% CI: 0.771-0.899) in the validation set, indicating strong predictive performance. Bootstrap internal validation demonstrated good agreement between predicted and observed outcomes in calibration curves. Decision curve analysis further indicated favorable clinical applicability of the nomogram. CONCLUSION: BMI, WC, TC, ALT, MetS-defined dysglycemia, and education level are independent risk factors for MAFLD. A dynamic nomogram for predicting MAFLD risk in patients with MetS was successfully developed and validated.

Humans

Delphi study robot consenso: Strategies for the implementation of robotic surgery in general surgery in the Spanish hospital network.

INTRODUCTION: The implementation of robotic surgery in public hospitals presents multiple logistical, educational, and organizational challenges. In the absence of unified guidelines, a national consensus is required to optimize its safe and efficient adoption. This study aimed to establish a set of consensus-based and measurable recommendations for the implementation of robotic surgery programs in hospitals within the Spanish National Health System, based on the experience of centres with established robotic programs and intended to serve as guidance for hospitals that are initiating or planning their implementation. METHODS: A national Delphi study was conducted with the participation of robotic surgery experts from 26 public hospitals. The expert panel was composed exclusively of digestive surgeons with experience in robotic surgery. Three iterative rounds of expert panel evaluation were conducted between March 2024 and March 2025. The questions were grouped into five thematic blocks. Consensus was defined as an agreement level of &#x2265;66.7%. Kendall's W coefficient was used to assess concordance. RESULTS: High levels of consensus were achieved on key aspects related to infrastructure, structured training, cost evaluation, and quality assurance mechanisms. Areas of disagreement were also identified, such as the need for a dedicated anaesthesiologist, purchase of accessory instruments during the initial phase, and official accreditation pathways. CONCLUSIONS: This study provides a guideline for developing a national robotic surgery strategy focused on patient safety, program sustainability, and standardized training of surgical teams. These recommendations can guide hospitals at different stages of robotic technology adoption. Given that the consensus was reached from an exclusively surgical perspective, the recommendations focus on patient safety, program sustainability, and standardized training of the surgical team, and should be interpreted in an adaptable manner according to each centre's context, case volume, and available resources.

Cirug&#xed;a Asistida por Robot

Pricing Combination Therapies: A Systematic Review of Value Attribution, Cost-Sharing Mechanisms and Policy Frameworks.

BACKGROUND: Combination therapies are increasingly central to modern pharmacotherapy, particularly in oncology and other high-burden diseases. However, pharmaceutical pricing and reimbursement systems remain largely designed for single-product-single-indication interventions. When multiple patented medicines are used together, especially when owned by different manufacturers, conventional pricing frameworks may struggle to align prices with the value of the combination while preserving incentives for innovation and timely patient access. OBJECTIVE: To identify, describe, and critically assess the methods, models, and policy frameworks proposed in the literature to establish prices for combination therapies, with particular attention to value attribution mechanisms, cost-sharing arrangements between manufacturers, and budget impact considerations. METHODS: A systematic literature review was conducted in accordance with PRISMA guidelines and a pre-registered Open Science Framework protocol. Searches were performed in MEDLINE, Scopus, Web of Science, EconLit, CRD databases, and grey literature sources for publications up to July 2025. Eligible studies analysed pricing approaches, economic models, reimbursement mechanisms, or policy frameworks relevant to combination therapies, including more recent multi-indication pricing literature. Given the heterogeneity of the literature, findings were synthesized using a structured narrative and thematic approach. RESULTS: Sixty-nine studies met the inclusion criteria. The literature was dominated by conceptual and policy analyses, with relatively few empirical or implementation-oriented studies. Value attribution emerged as the central methodological challenge in pricing combination therapies. Several complementary approaches were proposed to operationalise value attribution, including adaptations of indication- or pathway-based pricing, manufacturer cost-sharing arrangements, managed entry agreements, and outcome-based reimbursement mechanisms. Empirical evidence suggests that health systems continue to rely primarily on pragmatic and often partial solutions rather than fully specified pricing frameworks. A complementary review of the multi-indication pricing literature indicates that, although the two fields address different pricing problems, they share important methodological and institutional lessons that can inform the development of pricing frameworks for combination therapies. CONCLUSIONS: The literature provides a growing repertoire of conceptual approaches for pricing combination therapies but limited empirical evidence on implementation. Pricing frameworks should place value attribution at their core while combining complementary policy mechanisms adapted to national pricing and reimbursement systems. Lessons from multi-indication pricing provide a valuable foundation but require additional governance mechanisms to address value attribution, multi-manufacturer negotiation, and implementation challenges specific to combination therapies.

Journal Article

Volumetric bone marrow cellularity (VBMC) assessment from routinely processed trephines using three-dimensional x-ray histology and gaussian peak modelling.

Objective.Bone marrow cellularity is routinely estimated from a small number of two-dimensional histology sections, making assessment sensitive to section representativeness, processing artefacts and observer interpretation. Three-dimensional (3D) x-ray histology (XRH), using x-ray computed microtomography (&#xb5;CT), enables non-destructive whole-block imaging of trephine biopsies. This study evaluated whether XRH combined with Gaussian peak modelling could provide a pragmatic whole-block volumetric bone marrow cellularity (VBMC) estimate from formalin-fixed paraffin-embedded (FFPE) trephine biopsy blocks.Approach.Six routinely processed FFPE bone marrow trephine blocks were imaged using &#xb5;CT-based XRH at &#x223c;15 &#xb5;m spatial resolution. VBMC was defined as the red-marrow (RM) fraction of the marrow soft-tissue compartment, RM/(RM + intra-biopsy wax), with wax serving as the volumetric proxy for adipocyte/yellow marrow space. Whole-volume greyscale histograms were modelled using a three-peak Gaussian approach representing intra-biopsy wax, RM and demineralised trabecular matrix. Peak-height and area-under-the-curve metrics were compared with whole-volume 3D segmentation and clinical two-dimensional (2D) cellularity estimates.Main Results.Gaussian peak modelling successfully approximated the segmented tissue-phase distributions. The peak-height-derived VBMC metric showed the closest agreement with whole-volume 3D segmentation, with an average absolute percentage difference of 9.3%, compared with 18.6% for clinical expert 2D cellularity estimates. The area-under-the-curve metric followed similar trends but consistently overestimated VBMC. Clinical 2D cellularity broadly followed whole-biopsy trends but showed one discordant case not explained by slice-position sampling alone. XRH also enabled unrestricted virtual reslicing and visualisation of sectioning-associated artefacts prior to further microtomy.Significance.Pre-sectioning XRH combined with Gaussian peak modelling provides a rapid, segmentation-free route to volumetric cellularity estimation from intact clinical FFPE trephine blocks. The approach supports objective whole-biopsy assessment while remaining compatible with routine histopathology workflows, reflecting the expected limitations of section-based visual estimation despite its role as the current clinical standard. In the near term, it could provide a non-disruptive adjunct to conventional 2D cellularity reporting, pending larger validation studies.

Imaging, Three-Dimensional