Search PubMedSearch

SEARCH · Search PubMed

Results for “interreviewer agreement”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

40 recordsLinked to original sources

Artificial Intelligence Cannot Replace Peer Reviewers but May Help Editors Triage: A Comparative Analysis of a Large Language Model and Human Reviewer Recommendations at the American Journal of Sports Medicine.

BACKGROUND: The peer review system faces increasing strain from rising manuscript volumes, reviewer fatigue, and well-documented interreviewer disagreement. Large language models (LLMs) have shown potential to support the peer review process, but their ability to replicate editorial decisions at high-impact medical journals and their utility as manuscript screening tools remain unknown. PURPOSE: To compare the agreement between an LLM and the final editorial decision on manuscripts submitted to the American Journal of Sports Medicine and to evaluate the potential of LLMs as a manuscript screening tool. STUDY DESIGN: Cross-sectional agreement study. METHODS: Fifty-four manuscripts randomly selected from submissions to the American Journal of Sports Medicine (September 2024-October 2024) were reviewed by a locally deployed LLM (Ministral 3 14B; Mistral AI) using a standardized prompt. The artificial intelligence (AI) produced a categorical recommendation (reject, cascade, revision, or accept) and a numerical score (0-100) for each manuscript. Agreement with the final editorial decision was assessed by Cohen kappa (4-category model) for pooled human reviewers (n = 139 reviews) and the AI (n = 54). Screening performance was evaluated by positive predictive value (PPV), sensitivity, and specificity. RESULTS: Pooled human reviewers demonstrated fair agreement with the final decision (&#x3ba; = 0.181 [P < .001]; 42.4% agreement), while the AI demonstrated slight, nonsignificant agreement (&#x3ba; = 0.126 [P = .099]; 37.0% agreement). The AI recommended revision for 61.1% of manuscripts, of which 72.7% were ultimately rejected or cascaded, demonstrating systematic "revision bias." When the AI recommended rejection, 54.5% of those manuscripts were ultimately rejected and 27.3% were cascaded; when the AI recommended cascade, 50% were rejected and 50% were cascaded. However, when the AI recommended rejection or cascade (n = 21), 90.5% received a final decision of rejection or cascade (PPV, 90.5%; specificity, 81.8%). Manuscripts with an AI score <70 were rejected or cascaded 88.0% of the time (PPV, 88.0%). CONCLUSION: AI cannot replicate the nuanced judgment of human peer reviewers at a high-impact sports medicine journal. When AI recommended rejection or cascade, 90.5% of manuscripts received that final decision (descriptive PPV, 90.5%; 95% CI, 71.1%-97.3%), suggesting potential utility as an exploratory first-pass screening tool warranting further validation in larger cohorts. However, AI could not reliably distinguish manuscripts destined for outright rejection from those that would be cascaded to a sister journal-an important limitation for editorial triage applications.

Sports Medicine

Sensor-based measures of knee brace adherence have low agreement with self-report methods: A multi-measure study among knee osteoarthritis patients.

OBJECTIVE: To explore agreement between self-report and objectively measured adherence to brace wearing by patients with knee osteoarthritis. METHOD: A single-arm observational analysis nested within the PROP OA randomised controlled trial (ISRCTN28555470). Of 237 adults with symptomatic knee osteoarthritis randomised to brace treatment, 60 were included in this sub-study investigating three different methods of assessing knee brace wear time over 26 weeks: 1. Self-report questionnaires (SRQ) at 12 weeks and 26 weeks; 2. Short message service (SMS) questions (days worn in past week, typical hours per day when worn) administered from week 1 to week 24; 3. A skin temperature sensor embedded in the brace, sampling every 10&#x202f;min for 26 weeks. The presence and reason for the sensor were concealed from participants. The estimated proportion of participants meeting "minimum brace use", defined a priori as &#x2265;1&#x202f;h on &#x2265;2 days in past week, was described for each measurement method, overall and by brace type (unloader, neutral). For temperature sensor measurements, time spent above 24&#xb0;C and time spent above 25&#xb0;C were used. Agreement between the measures was summarised by percentage agreement and kappa (&#x138;). RESULTS: The estimated proportions of participants meeting "minimum brace use" at 12 weeks were 83% (SRQ), 83% (SMS), 60% and 58% (temperature sensor, 24&#xb0;C and 25&#xb0;C thresholds, respectively). At 26 weeks, the corresponding estimates reduced to 72%, 71% (SMS at 24 weeks), 43% and 37%. Sensor data suggested the sharpest decline in brace use occurred within the first 12 weeks. Agreement between self-report measures was higher than between self-report measures and sensor (SRQ vs SMS at 12 weeks: 92% agreement, &#x138;=0.67 (95%CI: 0.34, 1.00); SRQ vs Sensor at 12 weeks: 74%, 0.35 (0.10, 0.60); SMS vs Sens at 12 weeks: 76%, 0.36 (0.05, 0.66). Agreement between all measurement methods reduced at 26 weeks. CONCLUSIONS: This novel use of a temperature sensor to monitor brace adherence in knee osteoarthritis indicates that self-report adherence substantially overestimates knee brace wearing time, with implications for clinical trials and practice.

Humans

Reliability, Device Agreement and Validity of Load-Velocity Profiles: A Systematic Review with Meta-analysis.

BACKGROUND: For a valid one-repetition maximum (1RM) prediction via load-velocity (LV) relationships, high reliability and accuracy must be assumed. OBJECTIVE: Since individual study results indicate ambivalent prediction, this systematic review and meta-analysis was designed to provide a updated and comprehensive overview, extending knowledge about the validity and reliability of commercially available velocity sensors in Part I and the validity and reliability of velocity-based 1RM prediction models in Part II. METHODS: A systematic literature search was conducted in PubMed/MEDLINE, Web of Science, and Scopus. Validity and/or reliability studies or velocity-based 1RM prediction evaluations were included. Methodological quality was assessed using adapted COSMIN. The analysis was performed for intraclass correlation coefficient (ICC), Lin's concordance correlation coefficient (CCC), and Pearson's correlation coefficient (r). The review was preregistered in PROSPERO (CRD42025634595). RESULTS: Sixty-three studies were included for sensor validity and reliability and 38 for 1RM prediction models. Part I: Velocity sensors demonstrated good-to-excellent pooled validity and device agreement (ICC&#x2009;=&#x2009;0.91-0.92 [0.83-0.97]; k&#x2009;=&#x2009;55 and 439, respectively); intra- and inter-day reliability were classified as good to excellent with ICC&#x2009;=&#x2009;0.90-0.91 [0.85-0.95] (k&#x2009;=&#x2009;228 and 608, respectively), with sensor technology moderating the results. However, substantial heterogeneity and wide ranges of study-level estimates indicated considerable variability across moderators, linear position transducer (LPT) generally showing more consistent performance than inertial measurement units (IMU). Part II: Velocity-based 1RM prediction showed ICCs&#x2009;=&#x2009;0.90 [0.83-0.94] (k&#x2009;=&#x2009;124) and ICC&#x2009;=&#x2009;0.91 [0.72-0.98] (k&#x2009;=&#x2009;9); for reliability and validity, respectively. DISCUSSION: Commercial velocity sensors generally provide high relative validity and reliability. Results varied depending on exercise complexity, intensity, sensor technology, and modeling approach. While velocity-based 1RM prediction demonstrated high average validity, large heterogeneity in lower body exercises significantly biased the results. Furthermore, the dearth of measurement error and agreement analyses prohibits final conclusions. CONCLUSION: Therefore, velocity-based monitoring and 1RM prediction require cautious interpretation, as sensor- and exercise-specific evidence remains limited.

Load&#x2013;velocity relationship

Endoscopic Ultrasound-Guided Versus Transjugular Portal Pressure Measurements: Systematic Review and Meta-Analysis.

PURPOSE: Published reviews of endoscopic ultrasound-guided portal pressure gradient (EUS-PPG) have emphasized feasibility and safety. We performed a systematic review and meta-analysis specifically to evaluate how closely EUS-based portal pressure measurements track invasive comparator measurements in prospective paired studies and to summarize agreement, technical success, and adverse events. METHODS: We searched major databases through January 2026 for prospective cohorts reporting same-patient EUS-based portal pressure measurement and invasive hemodynamic measurements. Correlations were pooled with random-effects models and analyzed separately for studies comparing EUS-PPG with hepatic venous pressure gradient (HVPG) and studies comparing EUS-based portal measurements with direct portal venous pressure. Agreement and threshold discordance were summarized descriptively. RESULTS: Six prospective cohorts (127 attempted procedures) were included. In studies using HVPG as the comparator, the pooled correlation was 0.82 (95% CI, 0.72-0.89; I2&#x2009;=&#x2009;0%). In studies comparing EUS-based portal measurements with direct portal venous pressure, the pooled correlation was 0.86 (95% CI, 0.72-0.93; I2&#x2009;=&#x2009;16.9%). Technical success was 95.3%. EUS-PPG-attributed adverse events occurred in 2.4% of procedures, with no procedure-related deaths. Agreement data were limited. Reported limits of agreement were wide (approximately -&#xa0;6 to&#x2009;+&#x2009;7&#xa0;mmHg), and discrepancies of 5&#xa0;mmHg or greater occurred in 4 of 30 paired measurements. CONCLUSIONS: EUS-based portal pressure measurement is feasible and shows a strong association with invasive hemodynamic comparators, but the evidence base remains small (six cohorts, 127 attempted procedures). Further study will be necessary to establish patient-level agreement, procedural reproducibility, EUS-specific clinically significant portal hypertension thresholds, and whether HVPG-based decision thresholds can be transferred to EUS-derived measurements.

Humans

Parents' Awareness of Their Adolescents' Sexual Behaviors and Experiences: A Concordance Analysis.

PURPOSE: Adolescence is a developmental period during which sexual exploration is normative. Parental awareness of adolescents' sexual behaviors can be protective during this stage; however, research suggests both parents and adolescents tend to avoid open conversations about these sensitive issues. This study aimed to examine concordance between parent proxy-reports and adolescent self-reports across three domains of adolescents' sexual behaviors and experiences. METHODS: Data were collected from a national sample of 522 parent-adolescent dyads (aged 15-17 years) from the AmeriSpeak panel (data collected May-September 2022). Prevalence-adjusted bias-adjusted kappa (PABAK) statistics were calculated to evaluate agreement. RESULTS: Findings revealed weak-to-moderate concordance regarding adolescents sending and receiving sexual photos (PABAKs = 0.54-0.79), whereas stronger concordance was observed for experiences related to nonconsensual sharing of sexual photos (PABAKs = 0.92-0.94), with high agreement largely driven by low reported prevalence of these experiences. Strong concordance regarding whether adolescents had ever had sex (PABAK = 0.80) was largely driven by both parents and adolescents reporting the adolescent had not had sex. Concordance regarding sexually transmitted infection and pregnancy prevention methods varied, ranging from weak to strong agreement across specific methods (or lack thereof; PABAKs = -0.07-0.79). DISCUSSION: Results highlight significant discrepancies in parents' knowledge about their adolescents' sexual behaviors and experiences. Interventions aimed at enhancing trust and open dialogue between parents and adolescents about sexuality-related topics, as well as ensuring adolescents have a variety of trusted sources to consult for sexual information, can help adolescents make informed sexual health decisions.

Humans

Quantitative Outcomes for Shared Assessment and Management in Forensic Mental Health: A Meta-Analysis and Systematic Review.

Despite leading models of mental health care encouraging user involvement, users in forensic mental health (FMH) report poor involvement given the difficulty in reconciling shared approaches with risk-averse and legally mandated settings. While previous research has demonstrated qualitative benefits to shared approaches in FMH and has led to a proliferation of self-rated assessment tools, there remains to quantify agreement on self-rated tools and to clarify the impact of shared approaches on care. This meta-analysis examines (1) the correlation between clinician and user ratings, (2) the predictive validity of self-ratings for violence, and (3) the effects of shared risk management on violence and restriction in FMH. Five databases were searched from inception to April 2024, selecting for adult FMH inpatients, shared risk assessment, needs assessment or violence management as interventions, and quantitative outcomes (correlation, agreement, predictive validity, and effect on violence or restriction rates). Fifteen quantitative evaluations were retained. One of three planned meta-analyses could be conducted, with seven records providing paired clinician-user t-tests. Eleven more records provided clinical recommendations on operationalizing shared approaches. Random-effects meta-analysis showed a significant and large paired standard difference of .95 (95% CI&#x2009;=&#x2009;[.49,1.42]) across tools, with significant differences in DUNDRUM-3, DUNDRUM-4, and CANFOR sub-models. While acknowledging between-study heterogeneity, results substantiate quantitative differences where clinicians generally rate more needs and lesser progress than users across tools, showing that self-ratings can and should be used to broach collaborative discussions on needs and progress during FMH treatment. There remains an evidence gap for quantitative benefits in care outcomes and a need to standardize agreement measures for future comparisons and clinical sub-group analyses.

Humans

Physical reconfiguration of limb electrodes for Precordial Bipolar Lead acquisition: Morphological validation against digital subtraction.

BACKGROUND: The V2&#xa0;-&#xa0;V1 Precordial Bipolar Lead (PBL) selectively evaluates the right-to-left retrosternal axis and has shown diagnostic value beyond the standard 12&#x2011;lead electrocardiogram. However, its use has been limited by the need for raw electrocardiographic data and post-processing software. This study evaluated whether a simple physical reconfiguration of limb electrodes could reproduce the digitally derived V2&#xa0;-&#xa0;V1 morphology with sufficient accuracy for clinical application. METHODS: Thirty-seven subjects underwent two sequential 10-s 12&#x2011;lead recordings using a Cardiovit FT-1 electrocardiograph sampled at 1000&#xa0;Hz. In the standard recording, the digital PBL was calculated as V2&#xa0;-&#xa0;V1. In the second recording, the right-arm and left-arm electrodes were repositioned to the V1 and V2 sites so that Lead I directly recorded the retrosternal dipole. Signals were filtered, synchronized, and analyzed using median beats. Morphological agreement was assessed with Pearson correlation on Z-normalized signals, while absolute agreement was evaluated using Lin's concordance correlation coefficient (CCC), intraclass correlation coefficient (ICC (Lewis, 1931; Nehb, 1938 [1,2])), root mean square error (RMSE), and Bland-Altman analysis. RESULTS: Mean Pearson correlation between digital and physical PBL was 0.955 (SD 0.043), with segment-specific correlations of 0.953 (SD 0.054) for QRS and 0.967 (SD 0.052) for ST-T. Lin's CCC and ICC(2,1) were both 0.871 (SD 0.110), and RMSE was 0.091 (SD 0.049) mV. Bland-Altman analysis showed minimal bias (-0.008&#xa0;mV). CONCLUSIONS: Physical acquisition of the V2&#xa0;-&#xa0;V1 PBL achieved high agreement with the digitally derived signal, supporting a simplified analog method for broader clinical implementation.

Humans

Performance of Photon-counting CT for Assessing Pretreatment Breast Cancer: Comparison with Mammography, MRI, and 18F-FDG PET/CT.

Background Photon-counting CT (PCCT) offers improved spatial resolution, contrast to noise ratio, and dose efficiency, but its clinical utility remains incompletely defined for breast cancer. Purpose To evaluate the feasibility of PCCT for pretreatment breast cancer assessment through comparisons with MRI, full-field digital mammography (FFDM), and fluorine 18 (18F) fluorodeoxyglucose (FDG) PET/CT. Materials and Methods In this prospective study (March-May 2025), female participants with breast lesions categorized as Breast Imaging Reporting and Data System 4C or higher at US or FFDM underwent breast MRI and multiphasic contrast-enhanced PCCT. 18F-FDG PET/CT was performed in a subset with locally advanced disease. Four radiologists independently evaluated lesion morphologic characteristics, additional findings, and clinical TNM stage. Agreement was analyzed using intraclass correlation coefficients (ICCs) and &#x3ba; statistics. The diagnostic performance for additional lesions and nodal metastasis was compared with the reference standard (pathologic examination). Results Among 126 participants (mean age, 58.1 years &#xb1; 12.3 [SD]), interreader agreement across PCCT, MRI, and FFDM was good to excellent. PCCT agreed with MRI for lesion characterization (&#x3ba; = 0.57-0.96) and clinical T categorization (&#x3ba; = 0.86-0.88), with highest agreement with pathologic size (ICC, 0.70-0.81). For 46 pathologically confirmed additional lesions, PCCT was more sensitive than FFDM (difference, 44% [95% CI: 19, 66]) and similar to MRI (difference, 7% [95% CI: -5, 21]). Additionally, 44% (95% CI: 27, 52) of microcalcifications were missed at PCCT versus FFDM. For pathologically confirmed nodal metastasis, PCCT was more sensitive (difference, 10% [95% CI: 1, 20]) and accurate (difference, 6% [95% CI: 1, 11]) than MRI. For clinical N category, PCCT agreed with PET/CT (&#x3ba; = 0.82 [95% CI: 0.62, 0.96]; n = 19). Two distant metastases identified at PCCT were consistent with 18F-FDG PET/CT and pathologic findings. Conclusion PCCT demonstrated similar performance to MRI for lesion characterization and detection of additional lesions, with better performance for nodal metastasis evaluation; however, detection of microcalcifications was limited. &#xa9; RSNA, 2026 Supplemental material is available for this article.

Humans

Which radiographic plane should be used to quantify the distal tibia angle on weightbearing CT images?

BACKGROUND: Precise quantification of distal tibial alignment is essential for planning corrective osteotomies and ankle joint replacement surgery. The lateral distal tibial angle (LDTA) is the principal radiographic parameter used for this purpose. While LDTA is increasingly measured on weightbearing cone-beam CT (WBCT) using two-dimensional coronal slices, the optimal measurement plane remains unclear. METHODS: In this retrospective comparative study, full-leg WBCT scans of patients scheduled for supramalleolar osteotomy (n&#x202f;=&#x202f;20; mean age 47&#x202f;&#xb1;&#x202f;12.8 years) were analyzed. LDTA was measured on three coronal planes of the distal tibial plafond (anterior edge, mid-dome, posterior edge) and compared with semi-automated three-dimensional (3D) tibial alignment measurements as the reference standard. RESULTS: Mid-dome LDTA showed no significant difference from the 3D reference (p&#x202f;>&#x202f;0.05) and demonstrated excellent agreement. Anterior measurements significantly overestimated LDTA, while posterior measurements underestimated it (both p&#x202f;<&#x202f;0.05), with only fair agreement. CONCLUSION: LDTA should be measured at the mid-dome of the distal tibial plafond on WBCT to ensure accurate and reproducible alignment assessment. LEVEL OF EVIDENCE: Level III - Retrospective Comparative Study.

Humans

Assessing the Concurrent Validity of the Australian Treatment Outcomes Profile in a Methamphetamine Dependent Treatment-Seeking Population.

INTRODUCTION: The Australian Treatment Outcomes Profile (ATOP) is a brief clinical tool assessing substance use, health and well-being used in Australian alcohol and other drug treatment services. It is validated for use with clients using alcohol, opioids and cannabis, but not yet for clients who primarily use methamphetamine. METHODS: An embedded validation study was undertaken in treatment-seeking adults enrolled in a randomised double-blind placebo-controlled trial of lisdexamfetamine for methamphetamine dependence with sites in New South Wales, South Australia and Victoria. Participant demographics were collected during study screening. The ATOP and comparators (Time Line Follow Back, Opiate Treatment Index, Depression Anxiety Stress Scale, WHOQOL-BREF and Personal Wellbeing Index) were collected at baseline. Continuous ATOP items were analysed using Pearson's correlation coefficient, and dichotomous items were analysed using Fleiss's &#x3ba;. Agreement was rated as strong where measures were &#x2265;&#x2009;0.50, moderate where agreement was 0.30-0.49, and weak where <&#x2009;0.30. RESULTS: One hundred and eighteen study participants (2018-2020) had data for concurrent validity analysis. Strong validity was demonstrated for physical health, psychological health, quality of life, injecting drug use and crime items, and for days of use for amphetamines, alcohol, cannabis and cocaine. There was weak validity for days of use for benzodiazepines. Heroin use days and other opioid use days were endorsed by fewer than five participants and were therefore unable to be assessed. DISCUSSION AND CONCLUSIONS: The ATOP is valid for use in a treatment-seeking methamphetamine-dependent population, expanding the range of tools for assessment and standardised outcome monitoring across different settings and services.

Humans

RR-interval-based atrial fibrillation detection and burden estimation: cross-dataset validation and calibration-aware probability analysis.

Objective.Atrial fibrillation (AF) burden has become an increasingly important endpoint in long-duration rhythm monitoring, but reliable burden estimation requires more than accurate AF detection alone. In particular, when burden is derived by aggregating predicted AF probabilities over time, probability calibration may directly affect burden validity under external dataset shift.Approach.This study developed an interpretable-interval feature model for AF detection and evaluated it using record-wise cross-validation on a development cohort and independent cross-dataset external validation on public Holter electrocardiographic databases. Window-level performance was assessed using the area under the receiver operating characteristic curve (ROC-AUC), area under the precision-recall curve (PR-AUC), Brier score, expected calibration error (ECE), and calibration intercept and calibration slope. Recording-level AF burden was estimated using both probability-based and hard-label aggregation and evaluated using mean absolute error (MAE) and agreement analyses.Main results.The model showed high discrimination in both development and external evaluation, with external ROC-AUC ofand PR-AUC of. However, external calibration deteriorated despite preserved ranking performance, with Brier score of, ECE(15) of, calibration intercept of, and calibration slope of. In the external cohort, probability-based burden estimation preserved strong association with reference burden but showed weaker raw agreement than hard-label aggregation, with MAE ofversus, consistent with systematic probability underprediction. Repeated external recalibration across record-level splits substantially improved probability quality and probability-based burden estimation. Median probability-burden MAE decreased fromwithout recalibration toafter Platt recalibration andafter isotonic recalibration, while median ECE(15) decreased fromtoand, respectively.Significance.These findings indicate that-interval-based AF detection maintained strong ranking performance in the tested external cohort, but probability calibration should be evaluated explicitly when predicted probabilities are aggregated into AF-burden estimates.

Atrial Fibrillation

Assay-dependent variability in peptide biomarker quantification: experimental evidence from renalase in chronic kidney disease.

BACKGROUND: Renalase is a promising biomarker for kidney disease, but published levels vary widely between studies. We hypothesised that variability in commercial enzyme-linked immunosorbent assays (ELISAs) kits and matrix effects (serum vs plasma) drive these inconsistencies. METHODS: Paired serum and plasma samples from 56 participants (28 chronic kidney disease (CKD) stages 2-5, 28 healthy controls) were tested using three commercial renalase ELISAs (BTLAB, Cloud-Clone, EIAab). We assessed intra-assay precision, inter-assay agreement (Spearman's rank correlation and Bland-Altman analysis on log10-transformed values), matrix effects, and associations with estimated glomerular filtration rate (eGFR). Diagnostic performance was evaluated by Receiver operating characteristic (ROC) analysis. RESULTS: Inter-assay renalase concentrations differed markedly (up to orders of magnitude), with weak inter-assay correlations (r&#x2009;&#x2264;&#x2009;0.25). Bland-Altman analyses revealed large, systematic biases between kits. Only the BTLAB assay showed consistent serum/plasma agreement, a significant correlation with eGFR (&#x3c1;&#x2009;&#x2248;&#x2009;0.32-0.42, p&#x2009;<&#x2009;0.05), and moderate discriminatory performance for CKD in serum (AUC = 0.70) and plasma (AUC = 0.68). Cloud-Clone and EIAab produced divergent results and strong matrix-dependent biases. CONCLUSIONS: Observed variability among commercial ELISA platforms may compromise comparability between studies. Harmonisation, standardised reference materials, and cross-validation are necessary before renalase assays can be used reliably in clinical practice.

Humans

Meniscal preservation in the age of biologics: toward a quantitative decision algorithm for personalized repair.

BACKGROUND: Despite advances in arthroscopic repair and biologic augmentation, surgical indication for meniscal tears remains heterogeneous. No standardized framework currently integrates biomechanical, clinical, and biological determinants to guide repair versus resection. PURPOSE: To develop a quantitative decision model-the Meniscal Preservation Score (MPS)-that unifies biomechanical and biological evidence to stratify reparability potential and standardize treatment selection in meniscal surgery. METHODS: A systematic evidence synthesis conducted in accordance with PRISMA 2020 reporting standards of studies published from 2000 to 2025 in PubMed, Embase, and Scopus identified key determinants of meniscal healing. Five consistent predictors-patient age, vascularity, tear morphology, associated pathology, and activity profile-were weighted through a two-round modified Delphi consensus among ten experienced knee surgeons. The resulting 0-9-point MPS was incorporated into a stepwise decision tree linking lesion morphology, biological context, and surgical strategy. Conceptual validation used 50 simulated cases and a retrospective cohort of 45 patients to test agreement between algorithm recommendations and expert surgical decisions. RESULTS: The MPS achieved 86% concordance with expert judgment in simulation and 84% agreement in clinical validation. In this retrospective exploratory cohort, cases in which surgical management was concordant with MPS recommendations demonstrated higher mean IKDC scores at 24&#xa0;months and lower observed reoperation rates. These findings should be interpreted as associative rather than causal, as treatment allocation was not controlled and discordant cases may have represented inherently more complex pathology. CONCLUSION: The MPS represents an evidence-informed decision-support framework designed to systematize reparability assessment. While exploratory analyses suggest structural coherence with expert reasoning, prospective implementation and external validation are required before clinical adoption as a predictive tool. LEVEL OF EVIDENCE: conceptual model with exploratory validation.

Humans

Potential Links Between Physiological and Perceptual Strain in High Heat Stress.

The physiological strain index (PhSI) is widely used to quantify thermoregulatory and cardiovascular strain during heat stress. However, direct physiological measurements may not always be feasible in occupational or athletic settings. Therefore, this study aimed to examine the relationship and agreement between perceptual strain and integrated physiological strain indices during high heat stress. Ten healthy, physically active, non-heat-acclimated males (29 (7) yr; 1.79 (0.11) m; 77.4 (9.3) kg) completed two randomized crossover exercise trials in hot-dry (HD) and warm-humid (WH) environments with equivalent wet-bulb globe temperatures. Heart rate, rectal temperature, skin temperature, rating of perceived exertion, and thermal sensation were measured at baseline and every 15&#xa0;minutes during 60&#xa0;minutes of cycling. Physiological strain index (PhSI), adaptive physiological strain index (aPhSI), and perceptual strain index (PeSI) were calculated using validated equations. Repeated-measures correlation analyses demonstrated very strong associations between PeSI and both PhSI and aPhSI under HD (Rrm&#xa0;=&#xa0;0.940-0.943) and WH (Rrm&#xa0;=&#xa0;0.980-0.982; all p&#xa0;<&#xa0;0.001). Receiver operating characteristic analyses demonstrated good-to-excellent discrimination of physiological strain by PeSI (AUC&#xa0;=&#xa0;0.895-0.969). Mixed-effects analyses showed that higher PeSI values were associated with increased PhSI (&#x3b2;&#xa0;=&#xa0;1.325, p&#xa0;<&#xa0;0.001) and aPhSI (&#x3b2;&#xa0;=&#xa0;1.459, p&#xa0;<&#xa0;0.001). However, Bland-Altman analyses demonstrated relatively small mean biases (0.4-0.9 AU) but wide limits of agreement (-2.8 to 4.2 AU), indicating that PeSI and physiological strain indices are not interchangeable. These findings suggest that PeSI may serve as a practical adjunctive or screening indicator of physiological strain when direct physiological measurements are unavailable.

Humans

Comparison between measured and synthesized posterior lead electrocardiograms during percutaneous coronary intervention-induced myocardial ischemia.

BACKGROUND: Posterior/inferolateral myocardial ischemia is frequently underrecognized on standard 12&#x2011;lead electrocardiography (ECG). Synthesized posterior leads derived from the standard 12&#x2011;lead ECG have been proposed as an alternative to directly measured posterior leads; however, their accuracy under controlled ischemic conditions has not been fully validated. METHODS: We prospectively enrolled 26 consecutive patients undergoing percutaneous coronary intervention (PCI) in whom simultaneously recorded measured and synthesized posterior lead ECGs (V7-V9) were obtained during balloon-induced myocardial ischemia. ST-segment deviation was measured at the ST junction (STJ), 40&#xa0;ms (ST1), and 80&#xa0;ms (ST2) thereafter. Agreement between measured and synthesized posterior leads was assessed using Pearson correlation and Bland-Altman analyses. As an exploratory patient-level analysis, diagnostic performance was compared with reciprocal anterior ST-segment depression (V1-V4). RESULTS: Strong correlations were observed between measured and synthesized posterior lead ST-segment deviations (V7: r&#xa0;=&#xa0;0.89; V8: r&#xa0;=&#xa0;0.86; V9: r&#xa0;=&#xa0;0.83; all P&#xa0;<&#xa0;0.001). Bland-Altman analysis demonstrated minimal systematic bias (within &#xb1;0.004&#xa0;mV) and narrow limits of agreement. Synthesized posterior leads showed higher diagnostic performance than reciprocal anterior ST-segment depression (AUC 0.917 vs. 0.708), although the difference was not statistically significant (DeLong test, P&#xa0;=&#xa0;0.197). Using a 0.05&#xa0;mV threshold, synthesized posterior leads demonstrated 83.3% sensitivity, 100% specificity, and 96.2% overall accuracy. CONCLUSIONS: Synthesized posterior leads closely reproduced measured posterior lead ST-segment deviations during percutaneous coronary intervention (PCI)-induced myocardial ischemia, supporting the technical validity of posterior lead reconstruction. Larger prospective studies are warranted to determine whether synthesized posterior leads provide incremental diagnostic value beyond careful interpretation of the standard 12&#x2011;lead ECG.

Humans

A prospective crossover study comparing ICCS-recommended and Palmer-adjusted filling rates in children with spina bifida.

OBJECTIVE: This study aimed to investigate whether the maximal cystometric capacity (MCC) in children with spina bifida (SB) is indeed lower, as predicted by the Palmer formula, and to evaluate the impact of different bladder filling rates on urodynamic parameters. MATERIALS AND METHODS: This prospective, randomized, two-sequence crossover-controlled study included 70 children aged 3-18 years with spina bifida under regular follow-up. In Group 1, the first two bladder fillings were performed at the ICCS-recommended rate, and the third at 75% of that rate (Palmer formula). In Group 2, the sequence was reversed. Urodynamic parameters, including maximal cystometric capacity (MCC), bladder compliance, detrusor activity, filling pressures, and detrusor leak point pressure (DLPP), were analyzed across fillings. RESULTS: Cystometric bladder capacity was lower during fillings performed at the Palmer-adjusted rate compared with those at the ICCS-recommended rate. The proportion of reduced compliance significantly decreased in Group 1 (p = 0.046) but remained unchanged in Group 2. A significant positive correlation was observed between expected bladder capacity (EBC) and measured MCC in both groups (&#x3c1; &#x2248; 0.5-0.6). The highest correlation and agreement were found in Group 1 during the third filling at the Palmer rate (ICC = 0.606). No significant intra- or intergroup differences were observed in detrusor pressure, end-filling pressure, DLPP, or overactive bladder prevalence. CONCLUSION: Bladder filling rate was associated with differences in both cystometric capacity and bladder compliance in children with spina bifida. Fillings performed according to the Palmer formula (approximately 75% of the ICCS-recommended rate) were associated with capacities that more closely approximated age-expected values and with modest differences in bladder compliance. Conversely, faster filling rates did not produce similar benefits. These findings suggest that slower filling strategies may improve measurement consistency and agreement with expected bladder capacity estimates. However, the magnitude and direct clinical impact of these differences should be interpreted cautiously, particularly in light of the potential influence of sequence-related effects.

Humans

A smartphone-integrated Pt@Cu-HCF nanozyme-based paper sensor for on-site determination of total antioxidant capacity in marine oils.

Total antioxidant capacity (TAC) serves as a key indicator for evaluating the nutritional quality of foods. In this study, we designed a platinum-embedded copper hexacyanoferrate (denoted as Pt@Cu-HCF) nanozyme that exhibits high oxidase-like activity, efficiently catalyzing the oxidation of chromogenic substrates to generate robust colorimetric signals. Antioxidants quench hydroxyl radicals (&#x2219;OH) produced during the catalytic process, leading to a concentration-dependent suppression of the color signal. Leveraging this mechanism, a smartphone-integrated, colorimetric paper sensor for on-site TAC quantification was developed, using vitamin E as the calibration standard. The sensor was applied to determine TAC in fish oil, algal oil, and krill oil, demonstrating a linear response range of 9.78-312.5&#xa0;&#x3bc;M and a limit of detection (LOD) of 6.41&#xa0;&#x3bc;M. Validation using real-world marine oil samples showed excellent agreement with a commercial assay kit, confirming the reliability and practical applicability of this portable sensor for TAC measurement in complex biological matrices.

Antioxidants

Bioprospecting microbial genomes to expand the biocatalytic toolbox of rubber oxygenases.

A set of rubber oxygenases was discovered through phylogenetic analysis and AI-based structural modeling of complexes of the putative enzymes with a substrate mimicking cis-1,4-polyisoprene. Sixteen candidate proteins were selected from thermophilic microorganisms, all sequence-related to the Latex clearing protein from Streptomyces sp. K30 (LcpK30). Sequence truncation and solubility tags were then evaluated to enhance protein expression, with the SUMO tag proving to be the most effective. Including LcpK30, nine heme-containing oxygenases were successfully expressed in E. coli NEB 10-beta cells, purified (35-157 mg L-1 yield) and characterized. Steady-state kinetics revealed significant rubber latex-degrading properties for six of them, with the truncated SUMO-fused LcpK30 (SUMO-LcpK30T) showing activity in agreement with literature. Notably, the catalytic efficiencies of all the expressed homologs lay within one order of magnitude and the oxygenase from Thermomonospora echinospora was found to be particularly promising in terms of activity, especially at high latex concentrations (more than 1% w/v). The analysis of reaction mixtures by both HPLC and HPLC-MS confirmed the oxidation of cis-1,4-polyisoprene to form the expected isoprenoid oligomers (n&#x202f;=&#x202f;2-12), whose distribution was consistent with the usual endo-type cleavage pattern in all but one case. This bioprospecting effort afforded a platform of new rubber-degrading enzymes with diverse efficiencies and product profiles, capable of adapting to targeted applications.

Oxygenases