Search PubMedSearch

SEARCH · Search PubMed

Results for “randomized”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,081 records · Page 60Linked to original sources

Prevalence of Slowly Expanding Lesions in Patients With Multiple Sclerosis: A Systematic Review and Meta-Analysis.

BACKGROUND AND OBJECTIVES: Chronic active lesions (CALs) reflect chronic inflammation in multiple sclerosis (MS). Slowly expanding lesions (SELs) are CALs identified on conventional MRI by linear, concentric expansion over time, while paramagnetic rim lesions (PRLs) are CALs characterized by a paramagnetic rim on susceptibility-sensitive MRI. However, the prevalence of SELs and their overlap with PRLs remain unclear. The aims of this study were to (1) estimate the proportion of SELs among all T2 lesions and the proportion of patients with at least 1 SEL and (2) assess the proportion of SELs overlapping with PRLs. METHODS: We systematically searched PubMed, Scopus, Web of Science, and Embase on February 1, 2026, for studies evaluating SELs in MS. At least 2 authors independently assessed study eligibility. Primary outcomes were the pooled proportion of SELs among T2 lesions and the proportion of patients with at least 1 SEL. We estimated mean per-patient volumes of SELs and total T2 lesions and the proportion of SELs overlapping with PRLs. Random-effects generalized linear mixed-effects models and inverse-variance methods were used, with between-study heterogeneity assessed using τ2 and I2 and robustness using sensitivity analyses. Univariable meta-regression explored heterogeneity. PROSPERO: CRD42024603778. RESULTS: Of 5,980 records, 20 studies comprising 4,786 patients with MS were included (mean age: 43.6 ± 6.3 years; 63.7% female). Sample sizes varied by outcome. SELs accounted for 14% (95% CI 10-21) of all T2 lesions, and 78% (67-85) of patients had at least 1 SEL. After sensitivity analysis, per-patient mean volumes were 1.42 mL (0.79-2.06) for SELs and 10.6 mL (8.53-12.67) for total T2 lesions. In total, 11% (6-20) of SELs overlapped with PRLs. In subgroup analyses, proportions of SELs were similar in relapsing-remitting and progressive MS (15%), but the proportion of patients with at least 1 SEL was higher in progressive MS. Between-study heterogeneity was high across analyses with no significant sources identified. DISCUSSION: Although SELs represent a minority of T2 lesions, most patients have at least 1 SEL and a subset overlaps with PRLs, suggesting a partial correspondence between these 2 imaging markers of chronic inflammatory activity. Limitations include possible publication bias, high unexplained heterogeneity, differences in SEL identification methods, and differences in MRI time point number/timing.

Humans

Does Inquiry-Based Learning Improve Students' Critical Thinking? A Meta-Analysis Accounting for Control Group Variations.

BACKGROUND: Inquiry learning is widely recognized, through empirical studies, as an appropriate instruction in enhancing students' critical thinking, yet the results were varied across context. The previous meta-analysis did not include the control group variations as a potential moderator and the studies subject domain was limited only to science subjects. Consequently, it is difficult to generalize the effectiveness of IBL in enhancing critical thinking. This meta-analysis aims to investigate whether inquiry learning is effective in improving the students' critical thinking skills and examine the moderating roles of each study characteristic. Methods The literature search applying the PRISMA protocol 2020 was conducted by utilizing SCOPUS, ERIC, and DOAJ databases. A total of 57 studies from 51 articles, published from 2015 to 2025, were synthesized using a random-effects model with standardized mean difference (SMD). RESULTS: The analysis revealed that IBL has a large and significant effect on enhancing students' critical thinking (g = 1.336; 95% CI [1.061, 1.611]). However, substantial heterogeneity was observed (I 2 = 92.09%), suggesting variability across contexts. Moderator analyses revealed that the main moderator, control group variations, was statistically significant in moderating the effectiveness of IBL (Qm = 5.21; p = .022). in contrast, subject domain (Qm = 1.43; p = .698), education level (Qm = 1.11; p = .774), and country ( Q m  = 3.33; p = .650), were insignificantly moderating the effectiveness of inquiry learning. CONCLUSIONS: The present meta-analysis highlighted that IBL is effective in improving students' critical thinking. However, the effectiveness of IBL was relative to the type of control group variations. Its effect on critical thinking was greater when compared with teacher-centered learning but smaller when compared with other student-centered learning.

Thinking

Application of musculoskeletal ultrasound in postoperative rehabilitation assessment and monitoring after rotator cuff repair: A systematic review.

BACKGROUND: The development of rehabilitation protocols after rotator cuff repair has long lacked objective benchmarks. Traditional time‑based regimens are limited by considerable inter‑individual variability and an increased risk of re‑tear. Musculoskeletal ultrasound allows dynamic assessment of tendon healing and muscle morphology, yet evidence for directly linking its use to rehabilitation decisions remains scarce. OBJECTIVE: To systematically synthesize the evidence on the use of musculoskeletal ultrasound monitoring to inform rehabilitation decision‑making after rotator cuff repair. METHODS: Following the Preferred Reporting Items for Systematic Reviews and Meta‑Analyses (PRISMA) guidelines, we searched PubMed, China National Knowledge Infrastructure (CNKI), and Wanfang Data from January 2020 to April 2026. Original studies were included if they involved patients who had undergone rotator cuff repair, used musculoskeletal ultrasound (including gray‑scale ultrasound, elastography, etc.) to evaluate the rotator cuff tendons or shoulder muscles, and reported at least one parameter related to rehabilitation decision-making or functional outcomes. RESULTS: Eleven studies were included. Shear wave velocity (SWV), cross‑sectional area (CSA), and echo intensity (EI) were the most frequently reported ultrasound parameters. Available evidence indicated that SWV increased progressively after surgery, with an overall increase of approximately 22% to 25% from one week to 12 months postoperatively. This dynamic trajectory may serve as a reference baseline for judging rehabilitation progress. An abnormally elevated SWV in the early postoperative period was associated with an increased risk of re‑tear, suggesting that a more conservative rehabilitation strategy should be adopted. Tendon stiffness measured at 12 weeks after surgery independently predicted long‑term return to sport. Regarding muscle parameters, changes in CSA and EI were positively correlated with shoulder function scores, and the combination of these two parameters effectively identified patients with rehabilitation bottlenecks. CONCLUSION: Musculoskeletal ultrasound parameters are associated with the initiation of active movement, adjustment of exercise load, prediction of return‑to‑sport prognosis, and identification of retear risk. Among these, SWV shows particular promise as an objective monitoring parameter for supporting rehabilitation assessment after rotator cuff repair. Future randomized controlled trials are needed to determine whether ultrasound-informed assessment can improve rehabilitation outcomes compared with traditional time-based regimens, and to establish standardized measurement protocols and clinically applicable reference values. Key findings of this review are summarized in S1 File.

Humans

The effect of antiretroviral therapy adherence on viral load suppression rate among people living with HIV in Ethiopia: A systematic review and meta-analysis.

BACKGROUND: Antiretroviral therapy (ART) adherence is a key determinant of viral load suppression among people living with HIV (PLHIV). In Ethiopia, evidence on the magnitude of ART adherence and its effect on virological outcomes remains fragmented. This systematic review and meta-analysis aimed to estimate the pooled prevalence of ART adherence and viral load suppression, and to measure the association between adherence and viral suppression among PLHIV in Ethiopia. METHODS: This systematic review and meta-analysis used the PRISMA checklist for systematic reviews and meta-analyses. The review protocol has been registered onPROSPERO:(CRD420251125899). PubMed, ScienceDirect, Scopus, Epistemonikos, and Google Scholar were searched. The quality of included articles has been evaluated with a Newcastle-Ottawa Scale (NOS), adapted for observational studies. A random-effects model using restricted maximum likelihood (REML) with Knapp-Hartung adjustment was used to estimate pooled prevalence and odds ratio. Heterogeneity was assessed using I2, τ2, and Cochran's Q test. RESULTS: A total of 39 studies were included in the final analysis. The pooled prevalence of good ART adherence was 79.4% (95% CI: 74.8%-83.4%), while the pooled viral load suppression rate was 77.5% (95% CI: 72.5%-81.8%). The pooled odds ratio showed that good ART adherence was strongly associated with viral load suppression (OR = 6.30, 95% CI: 4.84-8.19). Substantial heterogeneity was observed across studies for both adherence and viral suppression outcomes (I2 > 90%). CONCLUSIONS: ART adherence and viral load suppression among PLHIV in Ethiopia are relatively high but remain below global targets. Good adherence was significantly associated with virologic suppression, highlighting adherence as a critical modifiable factor for achieving optimal treatment outcomes. Strengthening adherence support interventions is essential to improve virological success and advance progress toward HIV epidemic control.

Humans

Prenatal exposure to particulate matter (PM) and autism spectrum disorder (ASD) among children: a systematic review and meta-analysis.

The global surge in Autism Spectrum Disorder (ASD) cases, coupled with evidence linking prenatal Particulate Matter (PM) exposure to developmental disruption, demands a comprehensive review to design targeted health interventions. This systematic review and meta-analysis aim to evaluate the strength and consistency of evidence linking prenatal PM exposure to ASD across studies, quantifying this relation to identify actionable environmental risk thresholds. This study employed PRISMA protocols to systematically extract and evaluate evidence from PubMed, Web of Science, Scopus, and ScienceDirect (2010-2024), and screened 4,013 articles to identify qualified case-control and cohort studies (n=29). Data synthesis employed random-effects modeling, accompanied by comprehensive assessment through I2 statistics, Q-tests, funnel plots, Duval and Tweedie's trim-and-fill analysis, and Egger's regression, to ensure validity. A meta-analysis of 16&#xa0;case-control studies revealed a 34&#x202f;% increased risk of ASD associated with prenatal PM exposure (pooled OR=1.34; 95&#x202f;% CI: 1.13-1.54), despite substantial between-study heterogeneity (I2=94.02&#x202f;%, p<0.001). Publication bias was not significant (Egger's test p value=0.114). Critical trimester-specific analysis uncovered that third-trimester exposure significantly increased ASD risk (OR=1.17; 95&#x202f;% CI: 1.01-1.34), while first-trimester (OR=1.02; 95&#x202f;% CI: 0.92-1.11; I2=49.18&#x202f;%, p<0.10) and second-trimester exposures (OR=1.13; 95&#x202f;% CI: 0.88-1.38; I2=92.59&#x202f;%, p<0.001) showed non-significant associations. This review identified prenatal and early life exposure to PM as a risk factor for ASD, indicating a trimester-specific vulnerability. It highlighted the necessity of focused air quality interventions and targeted guidance to reduce prenatal PM exposure to alleviate ASD risk during the critical-window.

Child

Smartphone Apps for Preventing Adolescent Health Problems Among Health Care Professionals: Systematic Search and Quality Assessment.

BACKGROUND: Health care professionals must consider multiple dimensions of prevention when consulting with adolescents. Identifying risky behaviors early in adolescence is crucial for reducing both morbidity and mortality. General practitioners are increasingly eager to incorporate digital tools for prevention into their consultations with adolescents; however, the relevance and clinical validity of these digital tools are not always established or well-known. Consequently, primary care professionals require guidance and support in selecting relevant mobile health (mHealth) tools. OBJECTIVE: The aim of this study is to identify relevant and useful digital apps to help primary care professionals detect at-risk adolescents across all recommended areas of prevention: orthopedics, mental health, substance abuse, risk behaviors, sexual health, vaccinations, social relationships, and nutrition. METHODS: A systematic review of smartphone apps, with an analysis of content quality, was carried out by 4 researchers using the PRISMA (Preferred Reporting Items for Systematic Reviews and Meta-Analyses) checklist. The App Store and Google Play Store platforms were surveyed. The inclusion criteria were as follows: free of charge, date of last update, availability in French or English, relevance of the preventive approach to adolescents, and scientific validation. Four health care professionals assessed the apps: 2 selected the apps relevant to health care professionals, then 3 analyzed these apps using the French version of the Mobile App Rating Scale (MARS-F). Intraclass correlation coefficient, model (2,1) (2-way random effects, absolute agreement, single measures); standard error of measurement; and mean absolute error were also calculated. RESULTS: A total of 976 apps were identified, 49 of which had disappeared from the platforms prior to analysis. Nine apps were retained. Seven (0.72%) were included after evaluation using the MARS-F: 2 on mental health and 5 on sexual health (including 3 on contraception only). The mean MARS-F interrater score ranged from 2.5/5 to 3.8/5. The global MARS-F score demonstrated a pooled SD of 0.60 and an intraclass correlation coefficient (2,1) of 0.0003, resulting in a calculated standard error of measurement of 0.60. The average discrepancy between raters was a mean absolute error of 0.53. CONCLUSIONS: No similar studies have been identified in the literature that specifically focus on mobile apps designed to support health care professionals in delivering preventive care to adolescents. Of the 8 areas of prevention identified as relevant for adolescents, only 3 are addressed by the apps validated through our methodology (5 focus on sexual health). Consequently, current apps are insufficient to support health care professionals in their overall preventive work with adolescents. Such a review should be conducted systematically prior to the development of any new tool to prevent duplication and channel creative efforts toward truly innovative digital solutions. Furthermore, a thorough analysis of relevant, recommended websites is essential, as these resources complement the use of mobile apps designed for health care professionals.

Humans

Extended Reality Interventions for Osteoarthritis of the Knee and Recovery After Total Knee Arthroplasty: Systematic Review and Meta-Analyses.

BACKGROUND: Nonpharmacologic interventions are important for treating knee pain due to osteoarthritis or after total knee arthroplasty (TKA), and extended reality (XR) technology may enhance treatments for these indications. OBJECTIVE: This systematic review aimed to evaluate XR interventions for pain due to knee osteoarthritis (KOA) or for recovery after TKA. METHODS: Databases were searched through May 2023 and updated in December 2025. Eligible trials evaluated XR interventions to treat KOA pain or after TKA. We classified interventions by depth of immersion and clinical mechanism. We used the Grading of Recommendations Assessment, Development, and Evaluation (GRADE) criteria to determine the certainty of evidence for prioritized outcomes. Meta-analyses were performed when &#x2265;3 studies evaluated similar comparisons, outcomes, and time points. RESULTS: Eligible trials addressed KOA (k=12) or recovery after TKA (k=9). Sample sizes ranged from 36 to 306 participants, and most studies had a follow-up of &#x2264;3 months. Nineteen studies assessed pain-related functioning and pain intensity, and 5 assessed adverse events (AEs). For KOA, 10 studies examined interactive digital rehabilitation (IDR), and 2 examined virtual reality (VR)-digitally augmented exercise (DAE). IDR for KOA may result in better pain-related functioning (low certainty of evidence [COE]; pooled standardized mean difference [SMD] -0.59, 95% CI -1.11 to -0.06; prediction interval [PI] -1.72 to 0.55; k=5) and lower pain intensity at 6-8 weeks (low COE; pooled SMD -0.46, 95% CI -0.92 to 0.00; PI -1.39 to 0.47; k=4). VR-DAE for KOA (k=2) produced inconsistent results (very low COE). For post-TKA studies, 5 examined IDR, 2 examined VR-DAE, 1 examined VR-distraction, and 1 examined VR-psychoeducation. Post-TKA IDR may result in better pain-related functioning (low [k=4] and moderate COE [k=1]) but little to no difference in pain intensity (low-moderate COE; pooled SMD at 3-4 months -0.12, 95% CI -0.75 to 0.52; PI -1.63 to 1.27; k=3). VR-psychoeducation probably results in lower pain at 4 weeks (moderate COE; k=1), and VR-distraction may result in 6 months (low COE; k=1), whereas VR-DAE produced mixed findings (k=2; very low COE). IDR was not associated with AEs, and VR may not be associated with AEs for KOA (high and low COE), though AE reporting was uncommon (k=5) and evidence was very uncertain for post-TKA. CONCLUSIONS: IDR may augment treatment for KOA and post-TKA recovery, and VR may benefit post-TKA rehabilitation. This review is the first to stratify by level of immersion, clinical mechanism, and follow-up duration and to systematically evaluate AEs. IDR may be ready for integration into KOA care, while use after TKA needs more evidence. Randomized controlled trials with implementation outcomes could determine how XR interventions can be used for KOA, whereas trials evaluating efficacy and AEs are needed before their use for post-TKA.

Humans

Compliance With Ecological Momentary Assessment Among Patients With Cancer: Systematic Review and Meta-Analysis.

BACKGROUND: Patients with cancer often experience substantial fluctuations in psychological states during disease management. Traditional research tools are limited in capturing these dynamic changes in real time, constraining clinicians' understanding of patients' true conditions. Ecological momentary assessment (EMA) enables high-frequency, real-time data collection, providing patient-reported data with greater ecological validity. However, the effectiveness of EMA studies critically depends on patient compliance, and reported compliance rates vary widely, with a lack of systematic quantitative synthesis. OBJECTIVE: This study aims to systematically review and quantitatively analyze compliance with EMA among patients with cancer, and to examine whether EMA design characteristics were associated with compliance. METHODS: Web of Science, PubMed, Embase, Cochrane Library, CINAHL, PsycINFO, CNKI, and Wanfang databases were searched for literature published up to April 30, 2026. Compliance was defined as completed prompts divided by delivered prompts. Single-group proportions were pooled using logit transformation and random-effects models with the Hartung-Knapp-Sidik-Jonkman adjustment. Prediction intervals were calculated to describe the expected distribution of compliance in future comparable settings. Subgroup analyses, univariable meta-regressions, leave-one-out sensitivity analyses, and tests for small-study effects were performed. Risk of bias was assessed using the Joanna Briggs Institute Critical Appraisal Checklist for Studies Reporting Prevalence Data, methodological reporting quality was assessed using a modified Checklist for Reporting EMA Studies, and certainty of evidence was evaluated using the Grading of Recommendations Assessment, Development, and Evaluation approach. RESULTS: Twenty-three studies involving 13,565 participants were included. The pooled compliance rate was 78.55% (95% CI 73.48%-82.87%), with a prediction interval of 48.59%-93.41%. Subgroup analyses identified no robust differences across study characteristics. Although study length showed a statistically significant subgroup test, the result was not stable after excluding singleton categories. Meta-regression analyses similarly found no significant linear associations for study length, prompts per day, items per prompt, or assessment window. Leave-one-out analyses showed that no single study drove the pooled estimate. Regarding the risk of bias, 2 studies were judged as low, while 21 were judged as moderate risk. Quality scores ranged from 6.5 to 9.0, and the certainty of evidence for the pooled compliance rate was rated as very low according to the Grading of Recommendations Assessment, Development, and Evaluation approach. CONCLUSIONS: Overall compliance with EMA among patients with cancer was moderate to high, suggesting that repeated real-world assessment may be feasible in oncology research settings. Nevertheless, the very high heterogeneity, wide prediction interval, and very low certainty of evidence indicate that compliance is context-dependent. The pooled estimate should therefore be interpreted as an approximate benchmark rather than a universal expected rate. Future oncology EMA studies should use standardized compliance denominators, report missing prompts transparently, and prospectively evaluate patient-centered design strategies that reduce burden while preserving data quality.

Humans

Diagnostic Performance of Machine Learning for Systemic Lupus Erythematosus: Systematic Review and Meta-Analysis.

BACKGROUND: Early and accurate diagnosis of systemic lupus erythematosus (SLE) and its organ involvement is essential. Previous reviews of machine learning (ML) in SLE combined heterogeneous tasks and validation strategies and may have overinterpreted model performance. OBJECTIVE: This study evaluated the diagnostic performance of ML and deep learning (DL) models for 3 clinically distinct SLE-related tasks: SLE classification or diagnosis, lupus nephritis (LN) diagnosis, and neuropsychiatric systemic lupus erythematosus (NPSLE) discrimination. We also assessed methodological quality and certainty of evidence. METHODS: PubMed, Embase, Cochrane Library, Web of Science, and IEEE Xplore were searched from January 2014 to April 2026. Eligible peer-reviewed diagnostic accuracy studies developed or validated ML or DL models for 1 of the 3 prespecified tasks, used an accepted reference standard, and provided data for a 2&#xd7;2 contingency table. Bivariate random-effects meta-analyses with the Hartung-Knapp-Sidik-Jonkman adjustment were used to pool sensitivity and specificity. We reported 95% prediction intervals (PIs), assessed risk of bias using the Quality Assessment of Diagnostic Accuracy Studies for Artificial Intelligence tool (QUADAS-AI; Viknesh Sounderajah [Imperial College London]), and evaluated certainty of evidence using the Grading of Recommendations Assessment, Development, and Evaluation framework for diagnostic test accuracy. RESULTS: Twenty-nine studies were included: 17 for SLE classification, 5 for LN diagnosis, and 7 for NPSLE discrimination. In the primary task-stratified analysis, pooled sensitivity was 0.91 (95% CI 0.86-0.94; 95% PI 0.56-0.99), and pooled specificity was 0.94 (95% CI 0.91-0.96; 95% PI 0.69-0.99), with low heterogeneity (I&#xb2;=23.9% and 22.9%, respectively). DL models showed a sensitivity of 0.93 and specificity of 0.95, compared with 0.88 and 0.94 for traditional ML models. Certainty of evidence was high for most analyses but low for LN diagnosis because of inconsistency and imprecision. All studies were retrospective, and only 9 of 29 (31%) performed independent external validation. Overall risk of bias was high or unclear in 22 of 29 (75.9%) studies. No study reported model calibration, decision-curve analysis, or net clinical benefit. CONCLUSIONS: ML models showed promising diagnostic accuracy across 3 distinct SLE-related tasks, but wide PIs, limited external validation, and pervasive risk of bias restrict conclusions about real-world generalizability. Prospective multicenter studies with standardized tasks and reference standards, independent external validation, and formal assessment of calibration and clinical utility are required before clinical implementation.

Humans

Performance of AI-Based Screening Tools for Obstructive Sleep Apnea Across Apnea-Hypopnea Index Thresholds: Systematic Review and Meta-Analysis.

BACKGROUND: Obstructive sleep apnea (OSA) is highly prevalent but remains substantially underdiagnosed. Polysomnography (PSG) is the reference standard, but its cost and limited availability constrain large-scale case identification. AI-based screening tools may support risk stratification and referral prioritization, but their diagnostic accuracy across apnea-hypopnea index (AHI) thresholds remains uncertain. OBJECTIVE: This review aimed to systematically evaluate the diagnostic accuracy of AI-based OSA screening tools at AHI thresholds of &#x2265;5, &#x2265;15, and &#x2265;30 events/hour, with emphasis on models using non-PSG-derived inputs. METHODS: PubMed, Embase, Scopus, and Web of Science were searched for studies published from January 1, 2016, to May 3, 2026. Eligible studies included adults evaluated for suspected OSA or recruited from population-based cohorts, assessed AI-based models intended or interpretable for OSA screening, risk prediction, or screening-oriented severity classification, used PSG as the reference standard, and reported sufficient data to construct or reconstruct 2&#xd7;2 contingency tables. Diagnostic accuracy was synthesized separately by AHI threshold and input source using bivariate random-effects models, with 95% CIs and prediction intervals (PIs). Risk of bias and certainty of evidence were assessed using QUADAS-2 (Quality Assessment of Diagnostic Accuracy Studies 2) and GRADE (Grading of Recommendations Assessment, Development, and Evaluation), respectively. RESULTS: A total of 60 studies were included, of which 47 contributed data to the meta-analysis. At AHI thresholds of &#x2265;5, &#x2265;15, and &#x2265;30 events/hour, pooled sensitivities were 0.94 (95% CI 0.92-0.96; 95% PI 0.71-0.99), 0.87 (95% CI 0.84-0.89; 95% PI 0.66-0.96), and 0.83 (95% CI 0.79-0.87; 95% PI 0.61-0.94), respectively; the corresponding specificities were 0.77 (95% CI 0.69-0.84; 95% PI 0.30-0.96), 0.81 (95% CI 0.75-0.85; 95% PI 0.39-0.96), and 0.91 (95% CI 0.87-0.94; 95% PI 0.55-0.99), respectively. The corresponding areas under the summary receiver operating characteristic curves were 0.943, 0.907, and 0.920. For non-PSG-derived tools, sensitivities were 0.92, 0.85, and 0.81, and specificities were 0.70, 0.74, and 0.85 at the 3 thresholds, respectively. For PSG-derived models, sensitivities were 0.96, 0.90, and 0.85, and specificities were 0.82, 0.88, and 0.96, respectively. Exploratory subgroup analyses suggested performance variation across selected study and model characteristics, including region, algorithmic framework, data source, and validation method. CONCLUSIONS: AI-based tools showed generally favorable screening performance for OSA across clinically relevant AHI thresholds, although wide PIs suggest variable performance across future comparable populations and settings. By synthesizing diagnostic accuracy across 3 AHI thresholds and distinguishing non-PSG-derived from PSG-derived models, this review extends previous broad or modality-specific reviews and offers a clinically interpretable, pathway-specific basis for linking model performance to intended use. The findings may clarify potential roles for non-PSG-derived tools in front-end screening and referral prioritization and for PSG-derived models in reduced-channel assessment and sleep-laboratory workflow support. Given substantial heterogeneity, limited external validation, and low or very low certainty of evidence, prospective validation is needed before routine implementation.

Humans

Comparative Efficacy of Different AI Systems for Polyp Detection by Size During Colonoscopy: Systematic Review and Network Meta-Analysis.

BACKGROUND: Colorectal cancer remains a leading cause of death despite being largely preventable through polypectomy. AI systems designed to enhance polyp detection during colonoscopy have shown promise, but the extent to which they improve detection of different-sized polyps remains unclear. OBJECTIVE: This study compared the size-stratified efficacy of AI-assisted colonoscopy vs standard colonoscopy using the Hartung-Knapp-Sidik-Jonkman (HKSJ) method, and generated exploratory rankings while acknowledging all cross-platform comparisons are indirect. METHODS: This systematic review and network meta-analysis (NMA) searched PubMed, Embase, Cochrane CENTRAL, and Web of Science from inception to July 25, 2026, supplemented by citation searching. We included randomized controlled trials (RCTs) comparing AI-assisted vs standard colonoscopy in adults (&#x2265;18 years of age), reporting mean polyp detection counts stratified by size (&#x2264;5 mm, 6-9 mm, and &#x2265;10 mm). Two reviewers screened studies, extracted data, and assessed risk of bias using the Cochrane Risk of Bias 2.0. We conducted frequentist NMA using the HKSJ method with restricted maximum likelihood estimation, calculated 95% prediction intervals (PIs), and assessed heterogeneity using I2 and &#x3c4;2. Certainty of evidence was rated using the GRADE (Grading of Recommendations Assessment, Development, and Evaluation) framework. RESULTS: A total of 13 RCTs (4156 participants) compared 8 AI systems to standard colonoscopy, forming a network without direct AI comparisons. For diminutive polyps (&#x2264;5 mm), AI showed a modest advantage (standardized mean difference [SMD] 0.21, 95% CI 0.07 to 0.35, 95% PI -1.12 to 1.54), but substantial heterogeneity (I2=86.6%) and wide PI crossing the null indicated high uncertainty. EndoScreener showed the most consistent evidence (SMD 0.36, 95% CI 0.18-0.54). For small and large polyps, effects were minimal (SMD 0.02, 95% CI -0.02 to 0.06, 95% PI -0.03 to 0.07; SMD 0.01, 95% CI 0.00-0.02, 95% PI -0.01 to 0.03). GRADE certainty was very low for diminutive polyps and low for small and large polyps. Sensitivity analysis excluding Tianjin YuJin did not materially change findings. CONCLUSIONS: AI may modestly enhance diminutive polyp detection, but effects on small and large polyps are minimal, with no platform superiority. Given very low to low certainty, findings are hypothesis-generating. This exploratory NMA provides size-stratified comparisons that can inform future head-to-head trial design. Unlike prior reviews aggregating all polyp sizes, we show the overall AI benefit is driven by diminutive polyp detection, providing a framework for targeted deployment-prioritizing AI for diminutive polyp screening, with limited value for larger lesions. Head-to-head trials are urgently needed. TRIAL REGISTRATION: PROSPERO International Prospective Register of Systematic Reviews CRD420251266932; https://www.crd.york.ac.uk/PROSPERO/view/CRD420251266932.

Colonoscopy

Longitudinal Repeated Protein Measurements in a Multiethnic Cohort Identify Novel Diabetes Biomarkers That Reveal Unique Disease Pathways.

There is up to a fourfold increase in diabetes biomarkers identified with longitudinal repeated versus single time point proteomic measurements. The increase in biomarkers identified with longitudinal repeated measurements is supported by a similar proportion being nominated as causal for type 2 diabetes with Mendelian randomization. Proteins unique to the longitudinal repeated analyses highlighted biological pathways (e.g., posttranslational protein modification and cellular structure and cycle regulation) that were distinct from pathways enriched among the shared proteins (e.g., small-molecule metabolic and catabolic processes). Longitudinal protein measurements identify additional novel disease biomarkers and disparate biological pathways compared with single measurement analyses.

Journal Article

Radiographic assessment and orthodontic intervention effects on orthodontically induced root resorption: a systematic review and meta-analysis of clinical trials.

The relative contribution of radiographic methods and characteristics of the orthodontic intervention to orthodontically induced root resorption (OIRR) remains unknown. The aims of this systematic review and meta-analysis were to (1) estimate the pooled OIRR effect across orthodontic intervention versus comparator contrasts, (2) compare pooled estimates by radiographic method (2D [two-dimensional] vs. 3D/CBCT [three-dimensional/cone-beam computed tomography]), and (3) explore whether force mechanics (intrusive versus nonintrusive) modified OIRR magnitude. Seven randomized controlled trials and one prospective study (January 2010-October 2025) were included. Only OIRR was the outcome, reported as correlation coefficients (r). The primary analysis combined within-study intervention-versus-comparator estimates. Subgroup analysis of 2D versus 3D/CBCT imaging was prespecified, whereas post-hoc analysis of intrusive versus nonintrusive mechanics was performed. The pooled analysis for the primary outcome showed a small, nonsignificant OIRR effect (r = 0.07; 95% confidence interval [CI]: -0.12 to 0.27; p = 0.372) with high heterogeneity (I2 = 84.0%). Radiographic method did not change the pooled estimates significantly (p = 0.331). Force-mechanics analysis showed that intrusive mechanics was related to significantly higher root resorption than nonintrusive mechanics (r = 0.40; 95% CI = 0.15 to 0.65 versus r = -0.03; 95% CI = -0.16 to 0.10; p < 0.001). This accounted for 87.1% of the between-study variance. The average orthodontic intervention effect on OIRR was small and not significant; however, the OIRR magnitude was strongly affected by force mechanics, particularly by intrusive forces. There was no significant difference in pooled estimates by radiographic method; however, 3D/CBCT provides superior volumetric quantification and should be used judiciously according ALARA (as low as reasonably achievable) principles.

Root Resorption

Genome-wide association study of estimated glomerular filtration rate using repeated measurements in the Taiwan Biobank.

BACKGROUND: Chronic kidney disease (CKD) is a major global public health issue, with genetic factors playing a significant role in kidney function. Although genome-wide association studies (GWAS) have identified numerous loci associated with estimated glomerular filtration rate (eGFR), most studies relied on a single time-point measurement, which limits the capacity to account for within-individual measurement variability. METHODS: We performed a repeated-measurement GWAS in the prospective Taiwan Biobank (Taiwanese ancestry; n = 25,004) using two repeated creatinine-based eGFR measurements. Repeated eGFR values were analyzed using a linear mixed-effects model with a subject-specific random intercept and time-varying covariates, providing a more precise estimate of eGFR level. Identified loci underwent functional annotation (expression quantitative trait locus, deleteriousness prediction, and epigenetic markers) and were compared with results from a single-measurement GWAS. RESULTS: Six loci associated with eGFR were identified, including four previously reported regions (1q22, 4q21.1, 11p14.1, and 17q21.2) and two additional loci (6p21.32 and 15q24.2). Functional annotation implicated several candidate genes-such as MUC1/EFNA1, SHROOM3, HLA-DQB1, MPPED2, NRG4, and PGAP3/FBXL20-in the regulation of kidney function. CONCLUSION: Incorporating repeated eGFR measurements into GWAS may improve phenotypic precision for identifying genetic associations with kidney function. This study identified eGFR-associated loci and biologically plausible candidate genes in a Taiwanese population, which require further replication and functional validation.

Chronic kidney disease

Spinal meningiomas: histopathological grading using a benchmark radiomics model with notes on disease control.

OBJECTIVE: Spinal meningiomas (SMs) are common primary spinal tumors for which surgery is considered the first-line treatment when safe and feasible. The ability to extrapolate the tumor grade from preoperative imaging may significantly inform early patient expectation-setting regarding recurrence. Building on radiomics studies in cranial meningiomas, the authors aimed to construct a benchmark radiomics model to preoperatively identify the histological grade of SMs. METHODS: Institutional surgical records from May 2012 to November 2025 were queried for pathology-confirmed meningiomas below the foramen magnum, with preoperative contrast-enhanced imaging available for segmentation. SMs were classified as low-grade (WHO grade 1) and high-grade (WHO grade 2 tumors and grade 1 tumors with atypia). Tumors were manually segmented, and features were extracted using the PyRadiomics software package. An ensemble model of k-nearest neighbors, random forest, and support vector machine classifiers was trained using nested cross-validation on a subset of 10 features to differentiate tumor grades. Clinical data for the cohort were also extracted, and disease control in an adjunctive clinical series was assessed. RESULTS: Seventy-four patients were included in radiomics analysis, with an area under the receiver operating characteristic curve of 0.879 and a mean F1 score of 0.748. The model's top 5 features were all texture features that differed significantly (p < 0.05) across low- and high-grade SMs. These included measures of tumor textural and contrast-enhancement heterogeneity, with overlap with features reported in radiomics models for histological grading of intracranial meningiomas. Fifty-five patients with a median radiographic follow-up of 22.2 (range 1.9-86.4) months remained for clinical analysis after exclusion of patients with less than 1 month of follow-up and syndromic meningiomas. Four recurrences occurred at a median of 20.8 (range 1.8-41.8) months. High-grade tumor pathology did not significantly impact progression-free survival (p = 0.682, log-rank test; Cox regression high vs low grade hazard ratio [HR] 0.62, 95% CI 0.06-6.11, p = 0.685). Subtotal resection was associated with poorer progression-free survival than gross-total resection (p = 0.004, log-rank test; Cox regression subtotal vs gross-total resection HR 10.62, 95% CI 1.46-77.05, p = 0.019). These findings remain contextualized within a relatively limited follow-up window and small recurrence event count, suggesting a need to characterize the interplay between tumor grade and extent of resection as drivers of local disease control in SMs. CONCLUSIONS: A preoperative radiomics model can stratify high-grade SMs using open-source tools applied to single-institution data.

Humans

Metabolomic Responses to Oral Glucose Tolerance Test and Hyperinsulinemic-euglycemic Clamp in CKD.

BACKGROUND: The oral glucose tolerance test (OGTT) captures integrated physiological responses involving intestinal glucose absorption, incretin signaling, and endogenous insulin secretion, whereas the hyperinsulinemic-euglycemic clamp (clamp) isolates insulin-mediated glucose uptake. Comparing plasma metabolomic responses to these two challenges may identify processes specific to intestinal nutrient delivery and how they vary in CKD. METHODS: Targeted plasma metabolomics was performed in 59 adults without diabetes (39 with CKD [eGFR <60 mL/min/1.73 m2] and 20 controls) from the Study of Glucose and Insulin in Renal Disease (SUGAR). Each participant underwent a 75-g OGTT and clamp approximately one week apart. Eighty-eight plasma metabolites were quantified at fasting and during each challenge. Metabolite levels were log-transformed and normalized using Systematic Error Removal Using Random Forest (SERRF). Metabolites were classified using adjusted regression slopes relating OGTT and clamp responses. RESULTS: The mean (SD) age and eGFR were 64 (13) years and 54 (26) mL/min/1.73 m2, respectively, and 41% were female. In the overall cohort, OGTT and clamp induced broad plasma metabolic changes, with 63 (72%) and 76 (86%) metabolites significantly altered from fasting, respectively. Seventy-three metabolites (83%) demonstrated a significant relationship between OGTT and clamp responses. Of these, 22 (25%) exhibited true concordance and 51 (58%) demonstrated similar directional changes but differed in magnitude. A total of 15 (17%) metabolites were discordant or non-corresponding, of which only three were discordant. The non-corresponding metabolites were enriched in amino acid metabolism. Eleven metabolites (13%) demonstrated differential responses between OGTT and clamp by CKD status, involving amino acid and glucose metabolism pathways. CONCLUSIONS: Metabolomic responses to OGTT and clamp were largely directionally concordant but differed in magnitude, with attenuation during OGTT. Discordant metabolites were rare, while non-corresponding metabolites were confined to amino acid pathways. CKD modified OGTT-clamp correspondence for metabolites involved in amino acid and glycolytic metabolism.

Journal Article

Long-term glycemic variability and risk of peripheral artery disease: a systematic review and meta-analysis of cohort studies.

BACKGROUND: A systematic review and meta-analysis to evaluate the impact of long-term glucose variability (GV) on the risk of developing peripheral artery disease (PAD). METHODS: The protocol was prospectively registered in PROSPERO (ID: CRD420251148763). Relevant longitudinal studies were identified through comprehensive searches of PubMed, Embase, and Web of Science. The primary outcome was the risk ratio (RR) of PAD comparing participants with high versus low GV. Summary effect sizes were calculated using a random-effects model to account for between-study heterogeneity. RESULTS: Eleven cohorts were included. Higher GV showed a positive association with PAD risk (RR: 1.42; 95% CI [1.21-1.66] p&#xa0;<&#xa0;0.001), although substantial heterogeneity was present (I 2&#xa0;=&#xa0;91%). This association was consistent across subgroups defined by region (Asian vs. Western), study design, diabetic status, GV metrics, PAD diagnostic methods, and adjustment for HbA1c (all p for subgroup differences > 0.05), except for follow-up duration. Studies with follow-up < 8 years showed a stronger association than those with &#x2265; 8 years (RR: 1.64 vs. 1.19; p for subgroup difference = 0.006). CONCLUSIONS: Elevated long-term GV appears to be associated with an increased risk of PAD. However, substantial heterogeneity across studies suggests that the magnitude of this association should be interpreted with caution.

Humans

Comparative Efficacy and Safety of Adjunctive 0.015% Triamcinolone Acetonide Topical Spray With Lokivetmab Versus Lokivetmab Monotherapy as 'Reactive Therapy' for Canine Atopic Dermatitis: A Randomised, Single-Blinded, Controlled Preliminary Trial.

BACKGROUND: Lokivetmab is considered an effective systemic therapy for dogs with atopic dermatitis (AD); however, utilisation of lokivetmab with adjunctive topical anti-inflammatory glucocorticoids has not been investigated in canine AD. OBJECTIVES: To evaluate the efficacy and safety of the combination therapy of lokivetmab and 0.015% triamcinolone acetonide (TCA) spray in dogs with AD. ANIMALS: Twenty dogs with nonseasonal AD. MATERIALS AND METHODS: This study was a randomised, single-blinded (investigators only) 4-week controlled trial. Dogs were randomised to receive lokivetmab monotherapy or lokivetmab-TCA spray per manufacturer instructions. Clinical assessments included the Canine Atopic Dermatitis Extent and Severity Index, 4th iteration (CADESI-04), CADESI-04E (i.e., extracted erythema grade alone) and the 10-grade pruritus Visual Analog Scale (PVAS10) at Day (D)0, D14 and D28. Complete blood count and serum biochemical analysis were performed on D0 and D28. RESULTS: The median CADESI-04 and CADESI-04E scores in the lokivetmab-TCA group were significantly reduced on D14 (p&#x2009;<&#x2009;0.05 for both) and D28 (p&#x2009;<&#x2009;0.01 for both) compared to baseline; there were no significant differences in lesional scores at any time point for the lokivetmab monotherapy group. Lokivetmab-TCA significantly reduced CADESI-04E values compared to the lokivetmab monotherapy group on D28 (p&#x2009;=&#x2009;0.04). Both interventions significantly reduced PVAS10 scores on D14 and D28 compared to D0 (p&#x2009;<&#x2009;0.05 for both groups). CONCLUSIONS AND CLINICAL RELEVANCE: Topical application of TCA spray may be a useful and safe adjunctive therapy for systemic lokivetmab to alleviate pruritus and clinical lesions in the initial management of canine AD patients.

Animals