Search PubMedSearch

SEARCH · Search PubMed

Results for “measurement error”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

670 records · Page 2Linked to original sources

Pilot Distractions and Interruptions in Airlines: Ranking of Sources by Analytic Hierarchy Process.

ObjectiveThis work establishes a methodological framework for sources of pilot distraction and interruptions in a structured model that can be used as a tool for cockpit design/procedure assessment.BackgroundPilots must complete complex tasks, and distractions can impair performance and lead to errors that can cause aircraft accidents. Although various cockpit distractors are examined individually, there is no integrated approach.MethodDistraction and interruption sources were identified through a literature review and confirmed/extended by interviews with airline pilots. Associated weights were determined through pairwise comparisons, yielding a hierarchical model using the Analytic Hierarchy Process.Results26 sources of pilot distraction and interruptions were quantified and categorized into four categories: communication, head-down time, responding to abnormal conditions & unexpected situations, and searching for traffic.ConclusionA taxonomic structure for assessment is achieved with the top 5 sources identified as communications, technical interruptions, experience in type, environmental factors, operational irregularities, and airspace high terrain, accounting for 63.07%.ApplicationThe structured system is a flexible assessment scale that provides a taxonomic framework for airline risk management, supports future research, and cockpit design efforts.

Humans

Civil liability in oral & maxillofacial surgery: Α systematic review.

Maxillofacial surgery is a surgical specialty with anatomical, functional, and aesthetic requirements, which makes it a field at increased risk of medical negligence and subsequent legal claims. The review was conducted in accordance with the PRISMA guidelines and used international medical and legal databases. Studies that analyzed court decisions, insurance claims, or recorded compensation data related to maxillofacial surgery were included. The extracted data included, among others, the year and country of publication, the causes of action, and the amounts of financial compensation. Most lawsuits originate from countries with developed medical liability systems, primarily the United States and the United Kingdom, and there has been an increasing trend in publications over the last two decades. The most common causes of lawsuits involve nerve injuries, delayed or incorrect diagnoses, technical errors during surgery, and inadequate informed consent. The amounts of compensation vary widely, from lower five-digit amounts in milder cases to particularly high amounts in cases of permanent functional or aesthetic damage, placing a significant overall financial burden on health systems. Medical negligence in maxillofacial surgery constitutes an important forensic and socioeconomic issue. Understanding the causes of lawsuits and their financial consequences can help improve clinical practice, inform patients, and prevent legal disputes.

Humans

Instruments for measuring body image in breast cancer patients: a systematic review of measurement properties.

PURPOSE: To evaluate the psychometric properties of PROMs for measuring body image in breast cancer patients. METHODS: In December 2024, a psychometric systematic review was performed in the nine databases. The COSMIN checklist was employed to evaluate the methodological quality and psychometric properties of the included body image measures. The level of evidence was assessed using the GRADE framework, and final recommendations were formulated for the scale. RESULTS: Thirty-eight articles evaluating fifteen PROMs were included in this review. Structural validity, internal consistency, and hypothesis testing had been most frequently evaluated. Measurement error had not been assessed for all PROMs. Twelve instruments show potential application value but require further research. The BAS-BC, PSPP, and ASI-R are not recommended for use, as these instruments do not meet the strict COSMIN thresholds for full recommendation. CONCLUSION: The BIS can be recommended as a temporary screening tool for assessing body image outcome in clinical practice. The BIRS can be tentatively advised for measuring specific postoperative body image changes. However, further comprehensive studies are required to validate the psychometric properties of existing PROMs.

Female

Premature closure underlies bias in medical diagnosis in students: A randomised controlled experiment.

OBJECTIVE: The purpose of the study reported in this article was to shed light on the cognitive mechanism mediating between biasing information and diagnostic error. The literature suggests at least two different hypotheses: premature closure leading biased participants to spend less time on diagnosis or increased competition between diagnostic hypotheses. The latter hypothesis predicts that biased participants would spend more time reaching a diagnosis. METHOD: Using the salient distracting findings (SDF) experimental paradigm, we biased 58 fourth-year medical students while diagnosing 12 clinical vignettes in a within-group incomplete block design under three conditions: cases presented without SDF, with SDF at the beginning and with SDF at the end. For each of these conditions, diagnostic accuracy, the number of SDF-related mistakes and time per word needed to process the case were recorded. The data were analysed using linear mixed modelling. Estimated marginal mean scores were reported. RESULTS: Participants confronted with salient distracting features (SDFs) at the beginning of a clinical case demonstrated significantly lower diagnostic accuracy (mean 0.11) compared with the No-SDF condition (0.27), representing a 61% reduction (F2,693&#x2009;=&#x2009;11.995, p&#x2009;<&#x2009;0.001), and made more SDF-related mistakes (F2, 693&#x2009;=&#x2009;16.395, p&#x2009;<&#x2009;0.001). When SDFs were presented at the end of the case, diagnostic accuracy was also reduced (mean 0.17; 36% reduction), but processing time did not differ from the No-SDF condition. Only early presentation of SDFs was associated with reduced processing time per word (F2,636&#x2009;=&#x2009;4.799, p&#x2009;<&#x2009;0.01), consistent with premature closure. CONCLUSION: These findings demonstrate that biasing information increases diagnostic error in medical students and that only early bias is associated with reduced information processing. The data do not support the competition hypothesis for early bias, as processing time did not increase under biasing conditions. Premature closure can therefore be directly observed rather than inferred, inviting further research.

Humans

The Impact of Video Game Experience on Surgical Performance: A Systematic Review.

OBJECTIVE: To assess whether video gaming experience is associated with improved surgical performance in the surgeon population across laparoscopic, robotic, and other surgical modalities, and evaluate its implications on surgical training and education. METHODS: A structured literature search was conducted across five databases on 29th November 2024 in adherence to PRISMA guidelines. Comparative studies evaluating surgical performance outcome data between surgeons with differing video game experience were eligible for inclusion. Outcomes evaluated were time to complete task, error rate, accuracy, economy of motion, and overall score. Included studies were assessed for risk of bias using ROBINS-I and RoB 2. RESULTS: 15 studies involving 641 participants were included, comprising one randomized controlled trial and 14 nonrandomized studies. All nonrandomized studies were judged to be at moderate or serious risk of bias, and the single randomized controlled trial was judged to be at serious risk of bias. Differences in study design and outcome measures meant that quantitative synthesis could not be performed. Across laparoscopic studies, video gaming experience was associated with improved overall score and time to task completion predominantly in the period prior to structured training interventions, with accuracy findings consistently favoring gamers in the two studies reporting this outcome. Error rate and economy of motion findings were inconsistent or predominantly nonsignificant. No meaningful association was identified across nonlaparoscopic modalities. CONCLUSIONS: Video gaming may be associated with improved surgical performance in surgeons, though this appears restricted to laparoscopic tasks and the pretraining intervention period. As a low-cost and accessible activity, video gaming may represent a practical informal adjunct to formal surgical training to help ease the transition into structured technical training. Surgical program directors need not alter existing selection criteria or training modules based on the available literature.

Video Games

Webcam-Based Real-Time Visual Feedback During Baduanjin Practice in Older Adults: 6-Week Pilot Randomized Study.

BACKGROUND: Baduanjin qigong is a traditional mind-body exercise used to support balance and physical health in older adults. Age-related changes in proprioception may make accurate self-directed performance difficult without external guidance. OBJECTIVE: The aim of this study is to explore whether webcam-based real-time visual feedback delivered during supervised laboratory sessions was associated with differences in webcam-derived 2D pose discrepancy and movement consistency during Baduanjin practice in older adults. METHODS: A total of 31 older adults were enrolled, and 28 participants with complete analyzable records were included in this complete-case dataset (feedback group, n=14; nonfeedback group, n=14). All sessions were conducted face-to-face in a supervised motion-analysis laboratory. Weekly 2D pose-discrepancy values were analyzed using a linear mixed-effects model with fixed effects for group, categorical week, and the group-by-week interaction and a participant-specific random intercept. Joint- and movement-specific participant-level 6-week means were analyzed exploratorily using Welch independent-samples t tests. Holm correction was applied across 24 exploratory contrasts (6 week-specific, 8 joint-specific, and 10 movement-specific comparisons), and Hedges g and 95% CIs were reported. Participant-specific weekly slopes and within-participant variability were additionally examined to directly assess longitudinal error drift. RESULTS: The linear mixed-effects model showed no significant group-by-week interaction (Wald &#x3c7;25=1.09; P=.96) and no significant overall week effect (Wald &#x3c7;25=6.40; P=.27). Averaged across 6 weeks, the feedback group had an estimated mean 2D pose discrepancy 1.20&#xb0; lower than the nonfeedback group (95% CI -2.38&#xb0; to -0.02&#xb0;; P=.046), although this marginal pilot finding was sensitive to an analytic approach. No week-specific contrast remained significant after Holm adjustment. Nominal right elbow, right shoulder, and right knee differences did not survive global Holm correction. Form 3 showed a lower mean discrepancy in the feedback group (mean difference -3.70&#xb0;, 95% CI -5.86&#xb0; to -1.54&#xb0;; Hedges g=-1.30; unadjusted P=.002; Holm-adjusted P=.04). Direct analyses of participant-specific slopes and within-participant SDs did not support a significant between-group difference in longitudinal error drift. CONCLUSIONS: In this small exploratory pilot study conducted under supervised laboratory conditions, the 6-week trajectories did not differ significantly between groups. A marginally lower average 2D pose discrepancy was observed in the feedback group across the 6 weeks, but no individual week- or joint-specific comparison remained significant after multiplicity adjustment. Form 3 was the only exploratory contrast that remained significant after global Holm correction. Direct longitudinal analyses did not demonstrate prevention of error drift. Larger studies using validated reference measurements, prespecified outcomes, and adequately powered longitudinal designs are required.

Humans

Proprioception Training and Surrogate Outcomes: A Systematic Review of Definitions, Measures, and Effectiveness Claims.

BACKGROUND: "Proprioception training" is widely advocated in rehabilitation and sports practice, yet the term encompasses heterogeneous constructs, interventions, and outcomes. Many trials infer proprioceptive benefits from surrogate outcomes (balance, strength, or pain) rather than direct psychophysical indices. OBJECTIVE: We aimed to examine how proprioception is defined and measured, and how improvement is claimed, in randomized controlled trials. METHODS: Following Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) 2020, PubMed, Scopus, and Web of Science were searched to October 2025. Eligible randomized controlled trials explicitly described interventions as "proprioceptive" or "sensorimotor training" and reported at least one proprioceptive outcome, either direct (e.g., joint position reproduction, threshold to detection of passive motion, active movement extent discrimination) or indirect (e.g., sway, balance). Methodological quality was appraised with the Physiotherapy Evidence Database (PEDro) scale and risk of bias using the Cochrane Risk of Bias 2 (RoB 2) tool. RESULTS: Fifty-one randomized controlled trials (n&#x2009;=&#x2009;2319) were included. Comparative synthesis showed that improvements inferred from surrogate outcomes were more frequent and often larger than improvements observed in direct psychophysical measures. Directly targeted practice, angle specific, attentionally demanding, and aligned with the measured proprioceptive submodality and task construct, produced the most consistent benefits in position-reproduction accuracy/error, movement-detection sensitivity, or discrimination performance, depending on the outcome assessed. In contrast, multimodal regimens (balance, strengthening, taping, manual therapy) commonly improved balance, pain, strength, or function without comparably consistent evidence of enhanced direct psychophysical proprioceptive function. CONCLUSIONS: Specific psychophysical components of proprioceptive function appear modifiable, but only when training explicitly targets the sensory construct measured. The field remains conceptually diffuse, with frequent conflation of sensorimotor performance and proprioception. Progress depends on defining proprioceptive submodalities a priori, privileging validated psychophysical outcomes over surrogate outcomes, and aligning intervention content with measurement to substantiate true perceptual learning rather than generic motor adaptation.

Journal Article

Temporal redistribution of control reveals age-related differences in task switching at the level of preparation.

Task-switching studies often report minimal age-related differences in switch costs, leading to the conclusion that switching-related control processes are relatively preserved in aging. However, this conclusion is based on paradigms that confound preparatory and execution processes. This study examined whether age-related differences in semantic task-set reconfiguration may be underestimated due to this confound. In Experiment 1 (36 young and 30 older adults), participants performed an externally paced task-switching paradigm without control over preparation. In Experiment 2 (28 young and 28 older adults), a self-paced paradigm allowed participants to initiate stimulus onset, enabling measurement of preparation time. Across both experiments, reaction time (RT) and error rate (ER) showed reliable age effects but no interactions between age and condition, whereas switching-related condition effects varied across measures and experiments. The expression of switching-related costs differed across measures and task structures. Local switch costs were expressed in ER in Experiment 1 but in RT in Experiment 2. Global switch costs (all-switch vs. all-repeat) were observed in execution measures only in Experiment 1. In Experiment 2, preparation time showed reliable mixing, local, and global switching effects, with age-related amplification emerging specifically for global switching. These findings indicate that switching-related costs are redistributed across processing stages and behavioral measures. The results suggest that age-related modulation of semantic task-set reconfiguration may emerge more clearly during preparation than task execution, particularly under continuous switching demands. Preparation time is interpreted cautiously as reflecting participant-regulated preparatory processes rather than a pure measure of preparation efficiency.

Humans

Reliability, Device Agreement and Validity of Load-Velocity Profiles: A Systematic Review with Meta-analysis.

BACKGROUND: For a valid one-repetition maximum (1RM) prediction via load-velocity (LV) relationships, high reliability and accuracy must be assumed. OBJECTIVE: Since individual study results indicate ambivalent prediction, this systematic review and meta-analysis was designed to provide a updated and comprehensive overview, extending knowledge about the validity and reliability of commercially available velocity sensors in Part I and the validity and reliability of velocity-based 1RM prediction models in Part II. METHODS: A systematic literature search was conducted in PubMed/MEDLINE, Web of Science, and Scopus. Validity and/or reliability studies or velocity-based 1RM prediction evaluations were included. Methodological quality was assessed using adapted COSMIN. The analysis was performed for intraclass correlation coefficient (ICC), Lin's concordance correlation coefficient (CCC), and Pearson's correlation coefficient (r). The review was preregistered in PROSPERO (CRD42025634595). RESULTS: Sixty-three studies were included for sensor validity and reliability and 38 for 1RM prediction models. Part I: Velocity sensors demonstrated good-to-excellent pooled validity and device agreement (ICC&#x2009;=&#x2009;0.91-0.92 [0.83-0.97]; k&#x2009;=&#x2009;55 and 439, respectively); intra- and inter-day reliability were classified as good to excellent with ICC&#x2009;=&#x2009;0.90-0.91 [0.85-0.95] (k&#x2009;=&#x2009;228 and 608, respectively), with sensor technology moderating the results. However, substantial heterogeneity and wide ranges of study-level estimates indicated considerable variability across moderators, linear position transducer (LPT) generally showing more consistent performance than inertial measurement units (IMU). Part II: Velocity-based 1RM prediction showed ICCs&#x2009;=&#x2009;0.90 [0.83-0.94] (k&#x2009;=&#x2009;124) and ICC&#x2009;=&#x2009;0.91 [0.72-0.98] (k&#x2009;=&#x2009;9); for reliability and validity, respectively. DISCUSSION: Commercial velocity sensors generally provide high relative validity and reliability. Results varied depending on exercise complexity, intensity, sensor technology, and modeling approach. While velocity-based 1RM prediction demonstrated high average validity, large heterogeneity in lower body exercises significantly biased the results. Furthermore, the dearth of measurement error and agreement analyses prohibits final conclusions. CONCLUSION: Therefore, velocity-based monitoring and 1RM prediction require cautious interpretation, as sensor- and exercise-specific evidence remains limited.

Load&#x2013;velocity relationship

Comparison of black carbon measurements using filter-specific reference transmittance to those using lab blanks or an average of unloaded filters.

Filter-based optical techniques compare light transmission intensities between loaded (I) and unloaded (I0) filters as a measure of light-absorbing mass for subsequent estimations of equivalent black carbon (eBC). We analyzed 5,379 15&#x2009;mm Teflon filters from the Household Air Pollution Intervention Network (HAPIN) trial to assess the influence that different methods of I0 estimations have on eBC measures. We compared eBC measurements using filter-specific I0 values (Method 1) to those using three other methods of I0 estimation: the lab blank scan from a given session (Method 2), the average of all pre-sample filter scans (Method 3), and the average of all lab blank filter scans (Method 4). We assessed the agreement between Method 1 and the alternative methods using Bland-Altman analysis. We also assessed the relationship between Method 1 and the alternative methods across the complete measurement range and after stratifying exposure data into quartiles according to Method 1 eBC exposures. The mean (SD) personal eBC exposure for Method 1 was 7.8&#x2009;&#x3bc;g/m3 (5.9), and exposures ranged from 1.3 to 46.8&#x2009;&#x3bc;g/m3. Compared to Method 1, eBC using Methods 2, 3, and 4 were higher by 0.7&#x2009;&#x3bc;g/m3, 0.1&#x2009;&#x3bc;g/m3, and 0.7&#x2009;&#x3bc;g/m3, respectively. The performances of linear regression models between Method 1 and all other methods were moderate to strong (R2 range: 0.42-0.93) in the second, third, and fourth quartiles; however, the models in the first quartile (eBC range: 1.3-2.9&#x2009;&#x3bc;g/m3) performed poorly (R2&#x2009;=&#x2009;0.25-0.26), with error approximately 25% of the mean. Our findings suggest that, in most instances, conventional methods for obtaining I0 values can be used to sufficiently characterize eBC; however, analyzing filters before sampling adds appreciably to the accuracy of eBC estimations in lower concentration settings.Implications: Filter-based optical techniques compare light transmission intensities between loaded (I) and unloaded (I0) filters as a measure of light-absorbing mass for subsequent estimations of equivalent black carbon (eBC). To assess the influence that different methods of I0 estimations have on eBC measures, we analyzed 5,379 15 mm Teflon filters from the Household Air Pollution Intervention Network (HAPIN) trial. Our findings suggest that, in most instances, conventional methods for obtaining I0 values can be used to sufficiently characterize eBC; however, analyzing filters before sampling adds appreciably to the accuracy of eBC estimations in lower concentration settings.

Soot

Analyzing salinity tolerance in grass carp (Ctenopharyngodon idella): Insights from genome-wide association study and genomic selection.

Grass carp (Ctenopharyngodon idella) is one of the most widely cultured freshwater fish species globally. However, the expansion of its farming scale faces severe limitation owing to freshwater scarcity; therefore, the development of strains with greater salinity tolerance is key for expanding production using brackish water resources. To investigate the genetic basis of salinity tolerance in grass carp, a genome-wide association study (GWAS) was conducted using 200 individuals representing extreme phenotypes, namely salinity-tolerant and salinity-sensitive groups. In total, 17 single nucleotide polymorphisms (SNPs) related to salinity tolerance were detected, which were distributed across 11 chromosomes. Through gene annotation, 38 candidate genes were obtained from these loci. Enrichment analysis revealed these candidate genes are primarily implicated in key biological processes, including osmotic regulation, energy metabolism, and stress responses. Analyses of different SNP densities revealed that the 5&#xa0;K SNP density panel can balance prediction accuracy and computational efficiency. The BayesA model achieved the highest prediction accuracy under the GWAS_Evenly selection strategy, with substantial reductions in mean absolute error and mean square error. This study reveals the genetic mechanisms of salinity tolerance in grass carp, which might be optimized through genomic selection, and provides insights for selectively breeding new varieties with greater salinity tolerance.

Animals

Not just when, but how: An exploratory dual-control approach to video feedback in motor learning.

The present study provides exploratory evidence for a novel dual-control paradigm. It examines whether combining temporal over video feedback timing with learner-controlled interactive playback functions (pause, slow-motion, rewind) would enhance motor skill acquisition beyond temporal autonomy alone. Sixty-four novice adults were randomly assigned to one of four conditions: Full Control (self-controlled timing + interactive replay), Partial Control (self-controlled timing + non-interactive replay), Yoked Full Control (externally controlled timing + interactive replay), or Yoked Partial Control (externally controlled timing + non-interactive replay). Motor accuracy (Radial Error), movement consistency (Bivariate Variable Error), technical execution, and self-efficacy were assessed at pre-test, 24-h retention, and 72-h retention following two acquisition sessions on a dart-throwing task (120 trials total). The Full Control group demonstrated the greatest and most durable learning gains across all outcomes. The Group &#xd7; Time interaction was significant across all dependent variables (&#x3b7;2&#x209a; ranging from 0.140 to 0.234), with Full Control demonstrating superior retention at both 24 and 72&#xa0;h relative to other groups (though differences relative to Partial Control were more pronounced at 72-h retention). Critically, the Yoked Full Control group showed comparatively weaker outcomes despite access to the same interactive playback functions. These findings suggest that interactive video tools may be most useful when learners can regulate both when feedback is accessed and how it is inspected. Theoretical and practical implications for the design of learner-centered video feedback systems are discussed.

Humans

Improving insurance deduction identification: a hybrid artificial intelligence model using machine learning and expert systems.

PURPOSE: Financial challenges in healthcare systems worldwide, especially in low- and middle-income countries like Iran, have increased hospitals' reliance on insurance reimbursements. Unrecognized insurance deductions often cause severe financial shortages, making efficient deduction management crucial. This study aimed to design a hybrid intelligent system for identifying and predicting insurance deductions by combining machine learning and expert system frameworks. DESIGN/METHODOLOGY/APPROACH: A mixed-methods design was applied in four stages. First, a scoping review identified the causes and patterns of insurance deductions. Second, interviews with 15 insurance experts produced a validated checklist and a dataset from inpatient billing records. Third, using the CRISP-DM methodology, machine learning algorithms were developed and tested in SPSS Modeler alongside a fuzzy expert system developed in MATLAB. Finally, the model was validated using the holdout method. FINDINGS: Four categories of deduction drivers were identified: service provision, registration errors, document submission issues, and revenue conversion processes. The CHAID decision tree outperformed other algorithms with a 99% precision rate and the lowest Mean Absolute Error (9.43). A brief assessment of potential overfitting was conducted to ensure that the CHAID model's high accuracy was interpreted cautiously and supported by the validation results. The fuzzy expert system with validated rules was adaptable for deduction classification, especially for cases unsuitable for quantitative modeling. ORIGINALITY/VALUE: The hybrid model improves detection and prevention of deductions, offering actionable insights for hospital administrators, insurers, and policymakers. Its implementation can enhance hospital information systems, streamline claims processing, and optimize revenue management amid financial constraints.

Machine Learning

Plasma proteomics: considerations for preanalytical variability; a systematic review with narrative synthesis.

BACKGROUND: The plasma proteome (PP) is a dynamic system subject to pathology-associated changes and a focus for novel disease biomarker discovery. Disease-related PP research assumes protein concentrations in test specimens accurately reflect the in&#xa0;vivo milieu. However, measures to maintain the physicochemical integrity of the proteome before assay are often rudimentary, poorly described, or lacking standardisation in published studies. Contrastingly, in laboratory medicine, there is an expectation that errors in the so-called "preanalytical phase" (PAP) that impact patient results are understood, monitored, and mitigated against, while also being well described in research publications. There is therefore scope for good practice from laboratory medicine to inform PP research workflows. This review considers factors in the PAP which may impact the validity of PP results. CONTENT: A systematic review was conducted per PRISMA guidelines, limited to English-language peer-reviewed studies (2014-2024). Candidate studies were imported, screened, and managed using Covidence systematic review software. SUMMARY: 15 eligible studies were reviewed, covering many relevant processes. 11 studies reported statistically significant differences in PP due to factors in the PAP. Temperature and time-to-processing were the most commonly reported factors affecting the PP, with significant effects reported in 8 studies. OUTLOOK: PAP variability can significantly affect results in PP studies. Careful consideration of the effect of each stage of the PAP is needed when working with the PP. In multicenter studies, pre-defined and research question-specific sample processing workflows are essential for reducing PAP variability, which helps ensure the validity of PP studies.

Humans

Astigmatic vector outcomes after FS-LASIK versus SMILE for high myopic astigmatism: a single-center retrospective comparative cohort study without cyclotorsion compensation.

PURPOSE: To compare astigmatic correction vector outcomes between femtosecond laser-assisted in situ keratomileusis (FS-LASIK) and small-incision lenticule extraction (SMILE, also termed Keratorefractive Lenticule Extraction, KLEx) without intraoperative cyclotorsion compensation in patients with high myopic astigmatism (-&#x2009;2.00 to&#x2009;-&#x2009;3.75 D), and to clarify procedure-specific correction tendencies under this non-standardized alignment protocol. METHODS: This single-center retrospective comparative cohort study enrolled 155 eyes (one eye randomly selected per patient) that underwent FS-LASIK (80 eyes) or SMILE/KLEx (75 eyes) for high myopic astigmatism correction from January 2023 to July 2024 in Beijing Fenglian Jiayue Lige Clinic. Intraoperative cyclotorsion compensation was intentionally disabled to isolate inherent procedural astigmatism correction characteristics. Standardized Alpins vectorial analysis was performed at 3&#xa0;months and 12&#xa0;months postoperatively. PRIMARY ENDPOINT: 12-month Alpins correction index (CI). Multivariable propensity score adjustment was applied to mitigate confounding by clinical treatment selection bias. Statistical multiplicity control was implemented for secondary vector and visual outcomes. RESULTS: Baseline demographic, refractive, corneal and ocular biometric parameters were balanced between groups after propensity matching. No statistically significant intergroup differences were detected in uncorrected distance visual acuity (UDVA), corrected distance visual acuity (CDVA), residual cylinder, safety index or efficacy index at 3 and 12&#xa0;months (all P&#x2009;>&#x2009;0.05). Under the non-cyclotorsion-compensated protocol, significant intergroup differences were identified in the magnitude of surgically induced astigmatism (SIA), correction index (CI), and magnitude error (ME) at both follow-up timepoints (all P&#x2009;<&#x2009;0.0001). Target induced astigmatism (TIA), difference vector (DV), index of success (IOS), and angle error (AE) magnitudes were comparable between groups (all P&#x2009;>&#x2009;0.05). The vector mean axis of DV differed significantly between groups at 3 and 12&#xa0;months (Watson-Williams circular test, all P&#x2009;<&#x2009;0.0001). No reoperations were documented in clinic medical records for either cohort. No standardized dry eye questionnaires, tear film testing or corneal nerve density metrics were collected to quantify dry eye adverse events; only unstructured clinical notes were reviewed for complication screening. CONCLUSIONS: Under surgical alignment without cyclotorsion compensation, FS-LASIK and SMILE/KLEx both yielded acceptable visual and refractive safety/efficacy for high myopic astigmatism (-&#x2009;2.00 to&#x2009;-&#x2009;3.75 D) at 1-year follow-up, but demonstrated divergent astigmatism correction tendencies: FS-LASIK exhibited relative astigmatism overcorrection (vector mean DV:&#x2009;-&#x2009;0.35&#x2009;&#xb1;&#x2009;0.43 D&#x2009;&#xd7;&#x2009;91&#xb0;, CI&#x2009;>&#x2009;1), while SMILE/KLEx showed relative undercorrection (vector mean DV:&#x2009;-&#x2009;0.21&#x2009;&#xb1;&#x2009;0.53 D&#x2009;&#xd7;&#x2009;12&#xb0;, CI&#x2009;<&#x2009;1). These correction biases are specific to the study's manual limbal alignment protocol without cyclotorsion tracking and cannot be generalized to modern optimized surgical platforms equipped with automated cyclotorsion compensation. Residual refractive errors across both groups are likely multifactorial, including differential corneal stromal healing responses, divergent femtosecond/excimer laser tissue modification mechanisms, and uncorrected intraoperative ocular cyclotorsion.

Humans

Robust optimisation for photon radiotherapy: A scoping review of models, paradigms, and reporting.

BACKGROUND AND PURPOSE: Robust optimisation offers an alternative to conventional margin-based photon radiotherapy planning by explicitly modelling uncertainty, but practice is variable and not standardised. MATERIALS AND METHODS: A scoping review was conducted to map robust optimisation for photon external beam radiotherapy. Electronic searches of Scopus, PubMed and Google Scholar (2000-2025, English language) identified planning studies that incorporated modelled uncertainties into the optimisation process and reported at least one robustness-related outcome. Data were charted on clinical context, uncertainty models, optimisation paradigms, robustness metrics and evidence for clinical implementation. RESULTS: Seventy-one studies were included. Most investigated prostate, breast or lung cancer and used intensity-modulated radiotherapy or volumetric-modulated arc therapy in commercial or research treatment planning systems. Scenario-based worst-case (minimax) optimisation was the dominant paradigm in clinically oriented work, while chance-constrained, conditional value at-risk, distributionally robust and adaptive formulations were confined to small methodological series. Uncertainty modelling focused mainly on rigid set-up error; fewer studies incorporated respiratory motion, inter-fraction anatomical change, dose-calculation uncertainty or biological variation. Robustness was evaluated with diverse scenario-based dose-volume metrics, probabilistic coverage measures, composite robustness indices and, less often, biological endpoints. Direct clinical implementation reports were scarce. CONCLUSION: Robust photon planning is technically feasible and generally maintains or improves target coverage and organ sparing compared with margin-based planning. However, heterogeneity in uncertainty models, optimisation configuration and robustness reporting limits comparison and synthesis. Pragmatic minimum standards are proposed to support future consensus and wider clinical adoption.

Humans

A framework for delivering real-time, instrument-relative navigation in transoral robotic surgery.

Transoral robotic surgery (TORS) is a minimally invasive, inside-out technique that, compared with traditional open approaches, provides fewer post-operative complications, shorter hospital stays, and improved survival for early-stage head and neck cancer. However, TORS is limited by its steep learning curve and poor visualization of deep tumor margins. This randomized crossover study evaluated a surgical navigation system's potential to enhance accuracy and user experience with real-time, instrument-relative feedback. Seven Teflon beads (d&#x2009;=&#x2009;2.381&#xa0;mm) were embedded at the tongue base of a porcine pharynx-and-larynx model. Tongue blade compression and retraction were applied to the model to mimic intraoperative tissue deformation, reproducing the anatomical shifts that occur relative to preoperative imaging. Eight participants used the da Vinci Surgical system to localize the beads by placing pins under two conditions: (a) preoperative computed tomography with no navigation; (b) model-based visual navigation with quantitative instrument-to-target metrics. Surgical accuracy was determined by calculating the target localization error (TLE, pin-to-bead Euclidean distance) and the angular error (AE, pin axis trajectory to bead). Accounting for training level and bead depth, surgical navigation reduced TLE by 5.44&#xa0;mm (95% CI, 4.02-6.86&#xa0;mm; p&#x2009;=&#x2009;2.00e-11) and AE by 8.47 degrees (95% CI, 6.21-10.72 degrees; p&#x2009;=&#x2009;5.17e-11). Impressions of the system were generally favorable using a 5-point Likert survey and task duration (p&#x2009;=&#x2009;0.26) or cognitive workload via the NASA-Task Load Index (p&#x2009;=&#x2009;0.22) were not significantly affected. The navigation system demonstrated translational promise, offering improved target localization accuracy and more consistent performance across experience levels, two critical determinants of surgical quality in TORS.

Robotic Surgical Procedures

Robotic Needle Insertion for CT-guided Percutaneous Biopsy of Thoracoabdominal Lesions: A Prospective Multicenter Randomized Trial.

Purpose To compare safety and feasibility between a novel CT-guided robotic system and the conventional freehand technique for puncture biopsy of thoracoabdominal lesions. Materials and Methods In this prospective multicenter randomized trial, individuals with suspected lesions were enrolled between July 2023 and April 2024 across three university teaching hospitals and randomized to the robot-assisted group (n = 82) or the freehand group (n = 83). Procedure outcomes included the technical success rate, targeting error, number of CT scans and needle adjustments, puncture time, and complications. Descriptive and inferential statistics were calculated. Results A total of 165 participants (mean age, 60 years &#xb1; 10 [SD]; 83 male) were included. Compared with the freehand group, the robot-assisted group demonstrated a higher technical success rate (97.56% [80 of 82] vs 62.65% [52 of 83], P < .001), lower targeting error (mean Euclidean deviation: 1.7 mm &#xb1; 1.1 vs 4.5 mm &#xb1; 3.9, P < .001), and fewer CT scans (mean, 4.3 &#xb1; 1.9 vs 5.2 &#xb1; 2.3; P = .002) and needle adjustments (mean, 0.7 &#xb1; 0.7 vs 1.6 &#xb1; 1.6; P = .003). Despite differences in geometric precision, both groups achieved 100% (82 of 82 and 83 of 83) diagnostic yield. The median puncture time was comparable between groups (5.5 minutes &#xb1; 4.3 vs 4.8 minutes &#xb1; 7.0, P = .50). During lung biopsies, the robot-assisted approach yielded fewer complications compared with the freehand approach (4.88% [four of 82] vs 16.87% [14 of 83], P = .014). Conclusion Compared with the freehand approach, robot-assisted biopsy yielded greater precision and reduced adjustments and complications while demonstrating noninferior diagnostic efficacy and comparable duration. Keywords: Robotic Needle Insertion, Biopsy, Thoracoabdominal Lesions, Robot-assisted Biopsy, CT-guided Intervention, Percutaneous Needle Biopsy, Randomized Controlled Trial, Algorithm Development, CT, Clinical Testing, Interventional-Body, Biopsy/Needle Aspiration, Percutaneous, Thorax, Abdomen/GI, Liver, Lung, Kidney &#xa9;RSNA, 2026.

Humans