Search PubMedSearch

SEARCH · Search PubMed

Results for “Artificial intelligence”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

130 records · Page 4Linked to original sources

Enhanced fracture detection on radiographs with AI assistance for clinicians: a systematic review and meta-analysis.

BACKGROUND: Emergency radiographic interpretation for fractures is prone to missed or misdiagnoses. Artificial intelligence (AI) is expected to become a powerful tool to assist clinicians in fracture detection. PURPOSE: A systematic review and meta-analysis was performed to assess whether AI improves clinicians' ability to detect fractures on radiographs. MATERIALS AND METHODS: A literature search was conducted in PubMed, Web of Science, and Cochrane Library for studies published between January 1, 2010, and October 10, 2025. A meta-analysis of diagnostic accuracy studies was performed using a Summary Receiver Operating Characteristic (SROC) curve. The quality of included studies was assessed using the Quality Assessment of Diagnostic Accuracy Studies 2 (QUADAS-2) tool. Subgroup analysis and meta-regression were conducted to explore potential sources of heterogeneity. RESULTS: A total of 26 studies were included . The pooled sensitivity of clinicians increased from 77% (95% CI: 72-81) to 87% (95% CI: 83-90) with AI assistance, while the pooled specificity improved from 88% (95% CI: 85-90) to 92% (95% CI: 89-94). The corresponding AUC values were 0.90 (95% CI: 0.87-0.92) before and 0.95 (95% CI: 0.93-0.97) after AI assistance. Eight studies were rated as high risk of bias. Subgroup analysis and meta-regression identified potential sources of heterogeneity, including fracture location, AI model type, high risk of bias, and reference standards. CONCLUSION: AI assistance significantly improves clinicians' diagnostic performance in detecting fractures on radiographs for extremity and trunk fractures.

Humans

Systematic evaluation of one-dimensional-to-two-dimensional near-infrared spectroscopy transformations with deep learning for quantifying coconut sap adulteration.

Near-infrared (NIR) spectroscopy have limitations when combined with deep learning (DL) algorithms because they rely on low-dimensional datasets. Therefore, we investigated the potential of transforming one-dimensional (1D) NIR spectra into two-dimensional (2D) spectrograms using synchronous and asynchronous techniques and the continuous wavelet transform (CWT) and their effectiveness by integrating with DL for detecting adulteration in coconut sap. NIR spectra (12,500-4000 cm-1) were collected from binary mixtures (0%-100%;w/w). The performance of all DL (convolutional neural networks-CNN, AlexNet and ResNet) models was compared with that of partial least squares (PLS). The models were ranked in the mentioned order based on their performances: 2D-CWT > 2D-asynchronous > 2D-synchronous > 1D/2D-PLS. The important features of the best model can be explained and visualized using gradient-weighted-class-activation-mapping. The findings highlight that the 1D-to-2D NIR data transformation combined with DL is a highly robust approach because it addresses the feature representation gap in NIR data and effectively captures the spatial-spectral correlations.

Spectroscopy, Near-Infrared

ORBIT: Oncogenic Representation Learning via Bi-Prototype Contrastive Learning in Hyperbolic Space for cancer driver gene identification.

Accurate identification of cancer driver genes is crucial for precision oncology but remains challenging due to the complexity of integrating heterogeneous data and modeling dynamic biological systems. To address these limitations, we propose ORBIT (Oncogenic Representation Learning via Bi-Prototype Contrastive Learning in Hyperbolic Space). Our framework synergistically fuses multi-omics profiles with functional network data using a context-adaptive graph reweighting mechanism to capture cancer-specific dynamics. The model employs a bi-prototype contrastive learning strategy within hyperbolic space, which aligns gene representations around distinct driver and non-driver semantic anchors while preserving the intrinsic hierarchy of biological networks. Comprehensive evaluations demonstrate that ORBIT achieves highly competitive stability in pan-cancer analysis while consistently outperforming state-of-the-art methods in cancer-specific predictions. Furthermore, functional enrichment analysis confirms that the model effectively segregates core cancer pathways, and drug sensitivity profiling validates the clinical relevance of the identified drivers. By integrating hyperbolic geometry with context-adaptive learning, ORBIT offers a robust and interpretable paradigm for precision medicine. The source codes and datasets are publicly accessible at https://github.com/spcho-dev/ORBIT.

Humans

Assessment of atypical glandular cell interpretation in Pap tests using the Hologic Genius Digital Diagnostics System.

Atypical glandular cells (AGC) are a diagnostic challenge. The aim of this study was to evaluate the efficacy and diagnostic performance of AGC detection on the Hologic Genius Digital Diagnostics System (HGDDS). A retrospective analysis of 451 ThinPrep Pap cases was conducted, including 207 cases of AGC, 27 cases of high-grade squamous intraepithelial lesion (HSIL), 25 cases of low-grade squamous intraepithelial lesion (LSIL), and 192 benign cases. All AGC cases had follow-up histologic diagnoses, with 66 cases subsequently diagnosed as adenocarcinoma. The slides were randomized, scanned, and analyzed by the HGDDS. Patient age and HPV test results were provided to reviewers, an experienced cytologist, who screened the cases, followed by two cytopathologists who independently examined the cases on the HGDDS. Diagnostic concordance between the two cytopathologists indicated strong agreement (κ = 0.829). Sensitivity of AGC on Papanicolaou (Pap) tests for adenocarcinoma detection on HGDDS was 98.5% and 95.5%, respectively, comparable to the original ThinPrep interpretation (OTPI). Specificity for adenocarcinoma detection was significantly higher (84.6% and 85.6%) with the HGDDS than 27.7% with OTPI. Overall, the diagnostic performance for AGC/HSIL interpretation to detect CIN2/3/adenocarcinoma appeared to have improved with HGDDS compared with OTPI, particularly for specificity and positive predictive value (PPV). This is the first study evaluating AGC diagnosis using the HGDDS. The findings demonstrate that the sensitivity of adenocarcinoma detection as AGC on HGDDS is comparable to the ThinPrep Imaging System, but the specificity and PPV are improved. This suggests the potential of artificial intelligence to augment the performance of cervical cancer screening.

Humans

Clinical applications of digital twin technology in In Vitro Fertilisation.

BACKGROUND: Digital twin technology, originating from aerospace and manufacturing industries, has emerged as a transformative tool in healthcare. In vitro fertilisation (IVF) faces persistent challenges including suboptimal embryo selection, unpredictable treatment outcomes, and limited personalisation of protocols. Despite advances in assisted reproductive technology, existing literature exhibits fragmentation: artificial intelligence applications in embryo selection, ovarian stimulation, and endometrial assessment have been developed independently without systematic integration into comprehensive treatment frameworks. Digital twin technology offers unprecedented opportunities to create virtual replicas of biological systems, enabling real-time monitoring, predictive modelling, and personalised treatment strategies. AIM: This narrative review aims to critically examine the current applications of digital twin technology in IVF, evaluate its potential benefits and limitations, synthesize existing evidence into an integrative conceptual model, and identify future directions for implementation in reproductive medicine. METHOD: A comprehensive narrative review was conducted using PubMed, Scopus, Web of Science, and IEEE Xplore databases. A narrative review approach was selected over systematic review to accommodate the heterogeneity of evidence types in this emerging field, including theoretical frameworks, simulation studies, and proof-of-concept implementations that would be excluded from systematic reviews. Search terms included "digital twin," "IVF," "in vitro fertilisation," "assisted reproductive technology," "embryo selection," and "predictive modelling." Studies published between 2015 and 2025 were included, focusing on original research articles, systematic reviews, and proof-of-concept studies describing digital twin applications in reproductive medicine. RESULTS: Digital twin technology in IVF demonstrates significant potential across multiple domains including embryo development simulation, ovarian response prediction, endometrial receptivity modelling, and personalised stimulation protocols. Current applications integrate artificial intelligence, machine learning algorithms, time-lapse imaging, and omics data to create comprehensive virtual models. Early evidence suggests improvements in embryo selection accuracy, ovarian response prediction, and treatment protocol optimization, though large-scale randomized controlled trials remain limited. Implementation challenges include data integration complexity, computational requirements, regulatory considerations, and validation requirements. CONCLUSION: Digital twin technology represents a paradigm shift in IVF practice, offering personalised, predictive, and precision medicine approaches. This review synthesizes existing evidence to propose an integrative conceptual model for digital twin implementation across the IVF treatment spectrum, identifies critical knowledge gaps, and establishes research priorities to advance clinical translation. Despite current limitations, continued advancement promises improved success rates and patient outcomes.

Humans

Performance of AI-Based Screening Tools for Obstructive Sleep Apnea Across Apnea-Hypopnea Index Thresholds: Systematic Review and Meta-Analysis.

BACKGROUND: Obstructive sleep apnea (OSA) is highly prevalent but remains substantially underdiagnosed. Polysomnography (PSG) is the reference standard, but its cost and limited availability constrain large-scale case identification. AI-based screening tools may support risk stratification and referral prioritization, but their diagnostic accuracy across apnea-hypopnea index (AHI) thresholds remains uncertain. OBJECTIVE: This review aimed to systematically evaluate the diagnostic accuracy of AI-based OSA screening tools at AHI thresholds of ≥5, ≥15, and ≥30 events/hour, with emphasis on models using non-PSG-derived inputs. METHODS: PubMed, Embase, Scopus, and Web of Science were searched for studies published from January 1, 2016, to May 3, 2026. Eligible studies included adults evaluated for suspected OSA or recruited from population-based cohorts, assessed AI-based models intended or interpretable for OSA screening, risk prediction, or screening-oriented severity classification, used PSG as the reference standard, and reported sufficient data to construct or reconstruct 2×2 contingency tables. Diagnostic accuracy was synthesized separately by AHI threshold and input source using bivariate random-effects models, with 95% CIs and prediction intervals (PIs). Risk of bias and certainty of evidence were assessed using QUADAS-2 (Quality Assessment of Diagnostic Accuracy Studies 2) and GRADE (Grading of Recommendations Assessment, Development, and Evaluation), respectively. RESULTS: A total of 60 studies were included, of which 47 contributed data to the meta-analysis. At AHI thresholds of ≥5, ≥15, and ≥30 events/hour, pooled sensitivities were 0.94 (95% CI 0.92-0.96; 95% PI 0.71-0.99), 0.87 (95% CI 0.84-0.89; 95% PI 0.66-0.96), and 0.83 (95% CI 0.79-0.87; 95% PI 0.61-0.94), respectively; the corresponding specificities were 0.77 (95% CI 0.69-0.84; 95% PI 0.30-0.96), 0.81 (95% CI 0.75-0.85; 95% PI 0.39-0.96), and 0.91 (95% CI 0.87-0.94; 95% PI 0.55-0.99), respectively. The corresponding areas under the summary receiver operating characteristic curves were 0.943, 0.907, and 0.920. For non-PSG-derived tools, sensitivities were 0.92, 0.85, and 0.81, and specificities were 0.70, 0.74, and 0.85 at the 3 thresholds, respectively. For PSG-derived models, sensitivities were 0.96, 0.90, and 0.85, and specificities were 0.82, 0.88, and 0.96, respectively. Exploratory subgroup analyses suggested performance variation across selected study and model characteristics, including region, algorithmic framework, data source, and validation method. CONCLUSIONS: AI-based tools showed generally favorable screening performance for OSA across clinically relevant AHI thresholds, although wide PIs suggest variable performance across future comparable populations and settings. By synthesizing diagnostic accuracy across 3 AHI thresholds and distinguishing non-PSG-derived from PSG-derived models, this review extends previous broad or modality-specific reviews and offers a clinically interpretable, pathway-specific basis for linking model performance to intended use. The findings may clarify potential roles for non-PSG-derived tools in front-end screening and referral prioritization and for PSG-derived models in reduced-channel assessment and sleep-laboratory workflow support. Given substantial heterogeneity, limited external validation, and low or very low certainty of evidence, prospective validation is needed before routine implementation.

Humans

Applications of metal-organic frameworks in smart packaging for food freshness indication: a comprehensive review.

Smart packaging is extensively studied for its multifunctional capabilities in antimicrobial activity, preservation, and atmosphere modification. Recently emerged metal-organic frameworks (MOFs) freshness-indicating packaging becomes a key research direction in smart packaging owing to its distinctive functions and physicochemical properties. As multifunctional materials, the unique porous structure and tunable properties of MOFs provide a distinctive approach for developing food packaging applications dedicated to food freshness indication. Existing MOFs-based smart packaging still faces potential safety risks and technical challenges in practical applications, and there remains a lack of integrated discussion that combines synthesis strategies, packaging design, optimization, and safety assessment. This review elaborates on the application of MOFs in freshness-indicating smart packaging, focusing on diverse MOFs synthesis strategies, the formats of smart packaging, types of indicator signals, and qualitative/quantitative analytical methods. It also delves into the methodology concepts of MOFs-based smart packaging and evaluates MOFs safety in food packaging by addressing potential risks. Studies show that MOFs-based smart packaging achieves qualitative and semi-quantitative analysis of food freshness through multiple signal modalities such as visible color change, fluorescence, and photothermal effects. This review emphasizes that safe MOFs design is critically important and should comply with the overall migration limit of <10 mg/dm2 specified in Regulation (EC) No 1935/2004, lanthanide element limit of <0.05 mg/kg, and FDA threshold of 1.5 &#x3bc;g/person/day. Comprehensive safety assessment and intelligent sensing platforms will constitute pivotal directions for advancing MOFs-based smart packaging toward practical application.

Food Packaging

Community-driven advances in computational mass spectrometry: The perspective of EuBIC-MS members.

Advances in data acquisition, artificial intelligence, and integrative bioinformatics are driving the rapid evolution of computational mass spectrometry, and in turn, transforming modern proteomics, metabolomics, and lipidomics. These developments have greatly increased the scale and complexity of mass spectrometry data, underscoring the importance of evolving accurate, transparent, efficient and reproducible data processing workflows. Addressing these challenges requires collaborative innovation that brings together expertise in software engineering, statistics, and biology. The European Bioinformatics Community for Mass Spectrometry (EuBIC-MS), an initiative of the European Proteomics Association (EuPA), fosters a culture of open, community-driven development through its biennial Developers Meetings and Winter Schools. This commentary summarizes the scientific background and outcomes of the EuBIC-MS Developers Meeting 2025, which took place in Novacella, Italy. Three keynote presentations highlighted major frontiers in the field: deep proteome and phosphoproteome profiling, text mining for protein-protein interaction extraction, and scalable proteomics for AI-driven drug discovery. Seven community-selected hackathons addressed emerging challenges such as single-cell proteomics data analysis, FAIR metadata extraction, deep learning frameworks, R-Python interoperability, and DIA validation. Together, these efforts demonstrate the potential for scientific and technical innovation to arise from open collaboration, and highlight how community-driven initiatives can accelerate progress in computational mass spectrometry. SIGNIFICANCE: Modern proteomics increasingly depends on computational advances to translate complex, high-dimensional data into biological knowledge. The EuBIC-MS Developers Meeting 2025 exemplifies how community-driven collaboration can directly accelerate this process by bringing together experts from bioinformatics, statistics, and experimental proteomics to co-develop open, interoperable, and reproducible analytical tools. By fostering shared software frameworks, transparent benchmarking, and collaborative problem solving, the EuBIC-MS community helps ensure that technological innovation translates into reliable biological insights. This collaborative model strengthens the foundation for quantitative, system-level understanding of proteomes and establishes a sustainable path for integrating artificial intelligence and next-generation data acquisition into routine biological discovery. This commentary shows some current highlights in the field of computational mass spectrometry and community-based approaches undertaken during the most recent Developers Meeting to solve these challenges. The approaches discussed and initiated during the meeting - ranging from deep proteome profiling and phosphosite mapping to text mining, single-cell data analysis, and FAIR metadata extraction - address key bottlenecks that currently limit the biological interpretability and comparability of proteomics data.

Mass Spectrometry

From prediction to mechanism: Explainable AI uncovers plasma and CSF proteomic signatures of Alzheimer's disease.

Alzheimer's disease (AD) plasma and cerebrospinal fluid (CSF) proteomics can distinguish AD from cognitively normal controls, but the generalizability of machine learning performance and the recurrence of biological signals across datasets require cautious interpretation. We developed an explainable artificial intelligence framework spanning two fluids and four ADNI proteomic datasets, covering 2082 modality specific samples, all analysed internally within ADNI. Phase 1 analysed plasma using a 119 analyte NULISA and targeted UPENN panel (n&#xa0;=&#xa0;727; 216&#xa0;CE, 511 controls). Phase 2 extended the analysis to CSF using SOMAscan7k, TMT-MS and targeted SET2, with Elecsys A&#x3b2;42, A&#x3b2;40, total tau and p-tau181 as anchor biomarkers. Only SOMAscan was subject-independent relative to Phase 1 plasma; TMT-MS and SET2 overlapped with Phase 1 for 96.0% and 97.7% of subjects and therefore are not independent replication cohorts. Under subject-level splits with fold internal preprocessing, we compared Elastic Net, Explainable Boosting Machines and gradient boosted trees with SHAP-based explanations. Among the candidate pipelines, we selected the pipeline with the highest held-out test ROC AUC for each platform; the selected values were 0.927 in plasma and 0.954-0.973 across the three CSF datasets. Because the same held out test performance was used for pipeline selection and headline reporting, these are optimistically selected single-holdout estimates, not unbiased estimates of generalizable or clinical performance. Explanations identified five recurring biological axes within ADNI: cholinergic (ACHE), tau/14-3-3 (YWHAG, YWHAZ, YWHAB, YWHAE), neuro-axonal (NEFL, NEFH), microglial/complement (CHIT1, SMOC1, CHI3L1, C7, CFH) and synaptic (NPTXR, NPTX2, DLG4, SYT5, VSNL1, ELAVL2). CSF analyses showed synaptic vesicle-cycle enrichment (q&#xa0;=&#xa0;2&#xa0;&#xd7;&#xa0;10-6), and CSF YWHAG correlated strongly with total tau (&#x3c1;&#xa0;=&#xa0;0.87). Cross-fluid directional concordance was modest overall (54-57%) but increased to 73-80% among mapped analyte/protein rows reaching q&#xa0;<&#xa0;0.05 in CSF. These findings provide hypothesis-generating, internally supported evidence within ADNI. Independent external cohorts with locked pipelines are required to evaluate generalizable performance and biological reproducibility; the overlapping TMT-MS and SET2 analyses should not be interpreted as independent replication.

Alzheimer Disease

AI Health message intervention: The role of message customization and message source in breast cancer screening among women of color.

OBJECTIVES: To examine the effectiveness of breast cancer screening messages with varying levels of customization (generic, targeted, and tailored) and to compare AI-generated versus human-generated messages. METHODS: A between-subjects experimental design with a control condition was employed. Message content followed a standardized structure and varied by level of customization: generic, targeted (demographic-based), and tailored (perceived susceptibility- and barrier-based). Messages were developed by either the authors or GenAI (ChatGPT-4o). A total of 391 participants recruited via Prolific were randomly assigned to five groups (generic, targeted-human, targeted-AI, tailored-human, and tailored-AI). Self-efficacy, behavioral intentions, attitudes, and message believability were measured using different scales. RESULTS: Customized (tailoring and targeting) health messages performed comparably to generic messages in shaping positive health outcomes. GenAI-generated messages also produced outcomes comparable to those of human-generated messages under standardized conditions. Significant negative indirect effects through message believability for the human-tailored condition was found relative to the generic condition. CONCLUSIONS: GenAI may be a useful tool for developing and customizing scalable health messages. Its effectiveness depends not only on customization but also on maintaining message quality, including readability, clarity, coherence, naturalness, and credibility. PRACTICAL IMPLICATIONS: GenAI may support health practitioners in developing customized and scalable breast cancer messages. However, professional review remains necessary to ensure that the message is culturally appropriate, responsive to patient concerns, and suitable for use alongside patient-provider communication.

Humans

Diagnostic Performance of Machine Learning for Systemic Lupus Erythematosus: Systematic Review and Meta-Analysis.

BACKGROUND: Early and accurate diagnosis of systemic lupus erythematosus (SLE) and its organ involvement is essential. Previous reviews of machine learning (ML) in SLE combined heterogeneous tasks and validation strategies and may have overinterpreted model performance. OBJECTIVE: This study evaluated the diagnostic performance of ML and deep learning (DL) models for 3 clinically distinct SLE-related tasks: SLE classification or diagnosis, lupus nephritis (LN) diagnosis, and neuropsychiatric systemic lupus erythematosus (NPSLE) discrimination. We also assessed methodological quality and certainty of evidence. METHODS: PubMed, Embase, Cochrane Library, Web of Science, and IEEE Xplore were searched from January 2014 to April 2026. Eligible peer-reviewed diagnostic accuracy studies developed or validated ML or DL models for 1 of the 3 prespecified tasks, used an accepted reference standard, and provided data for a 2&#xd7;2 contingency table. Bivariate random-effects meta-analyses with the Hartung-Knapp-Sidik-Jonkman adjustment were used to pool sensitivity and specificity. We reported 95% prediction intervals (PIs), assessed risk of bias using the Quality Assessment of Diagnostic Accuracy Studies for Artificial Intelligence tool (QUADAS-AI; Viknesh Sounderajah [Imperial College London]), and evaluated certainty of evidence using the Grading of Recommendations Assessment, Development, and Evaluation framework for diagnostic test accuracy. RESULTS: Twenty-nine studies were included: 17 for SLE classification, 5 for LN diagnosis, and 7 for NPSLE discrimination. In the primary task-stratified analysis, pooled sensitivity was 0.91 (95% CI 0.86-0.94; 95% PI 0.56-0.99), and pooled specificity was 0.94 (95% CI 0.91-0.96; 95% PI 0.69-0.99), with low heterogeneity (I&#xb2;=23.9% and 22.9%, respectively). DL models showed a sensitivity of 0.93 and specificity of 0.95, compared with 0.88 and 0.94 for traditional ML models. Certainty of evidence was high for most analyses but low for LN diagnosis because of inconsistency and imprecision. All studies were retrospective, and only 9 of 29 (31%) performed independent external validation. Overall risk of bias was high or unclear in 22 of 29 (75.9%) studies. No study reported model calibration, decision-curve analysis, or net clinical benefit. CONCLUSIONS: ML models showed promising diagnostic accuracy across 3 distinct SLE-related tasks, but wide PIs, limited external validation, and pervasive risk of bias restrict conclusions about real-world generalizability. Prospective multicenter studies with standardized tasks and reference standards, independent external validation, and formal assessment of calibration and clinical utility are required before clinical implementation.

Humans

A multi-scale fusion model based on multi-phase contrast-enhanced CT for predicting pancreatic cancer resectability.

Purpose.Develop a multi-scale fusion model (MSFM) based on multi-phase contrast-enhanced computed tomography (CECT) to predict pancreatic cancer (PC) resectability, thereby assisting expert decision-making.Methods.This retrospective study enrolled 280 patients with PC from four institutions, which were randomly divided into a training cohort (202 patients) and an independent test cohort (78 patients). Three-phase CECT images (arterial, venous, and delayed phases) were used for modeling. The MSFM comprises two sub-networks: (1) a multi-phase fusion network for extracting cross-phase shared fusion features, (2) a phase-specific branch network for capturing phase-specific features; and a post-fusion strategy to generate the final predictive score by integrating the shared fusion features and three groups of phase-specific features. Additionally, a human-machine fusion deep learning model (HMfDL) was constructed by fusing the predictive score of the MSFM with expert assessments.Results.In the independent test, the MSFM achieved an AUC (area under the receiver operating characteristic curve) of 0.8385 (95% CI: 0.7521-0.9249), accuracy of 84.62%, sensitivity of 72.00%, and specificity of 90.57%. This performance outperformed single-phase models (AUC range: 0.7638-0.7781), two-phase models (AUC range: 0.7826-0.7864), and ten states-of-the-art classifiers (AUC range: 0.7404-0.7796). The HMfDL further improved the performance, reaching an AUC of 0.8626 (95% CI: 0.7853-0.9400), accuracy of 91.03%, sensitivity of 80.00%, and specificity of 96.23%. Notably, the HMfDL corrected 58.82% of misdiagnosis made by experts.Conclusions. The MSFM effectively fuses multi-phase CECT to enable highly accurate predictions of PC resectability, and provides valuable support for expert decision-making through HMfDL.

Humans

Effectiveness of an AI-based home exercise app for rehabilitation of rotator cuff-related shoulder pain: A randomized controlled trial.

BACKGROUND: Rotator cuff-related shoulder pain contributes to disability and healthcare use. Although therapeutic exercise is first-line treatment, limited supervision and adherence may reduce its effectiveness; digital rehabilitation with real-time feedback may address these limitations. OBJECTIVES: To evaluate the effectiveness of adding a digital rehabilitation program to standard physiotherapy on pain, function, fear-avoidance beliefs, and healthcare utilization. DESIGN: Single-center, assessor-blinded, randomized controlled trial with two parallel groups. METHOD: Forty-six adults (mean age 59 years) with rotator cuff-related shoulder pain were randomized to 12 weeks of conventional physiotherapy or physiotherapy plus an AI-based digital rehabilitation program using computer vision for real-time feedback and performance monitoring. Outcomes were assessed at baseline and at 2, 4, and 12 weeks. Pain intensity (NPRS) was primary outcome; secondary outcomes included upper limb function (QuickDASH), fear-avoidance beliefs (FABQ), and post-intervention healthcare utilization. Analyses followed an intention-to-treat approach. RESULTS: Pain reduction exceeded the MCID (1.3) at 4 and 12 weeks. Between-group differences favoured the intervention at Weeks 2 and 4 (MD -0.7; 95% CI -1.13 to -0.14 and MD -1.01; 95% CI -1.8 to -0.2, respectively). Upper limb function improved more at Week 4 (MD -7.3; 95% CI -12.3 to -2.2). FABQ scores decreased more at Week 12 (MD -7.6; 95% CI -14 to -0.5). Fewer participants in the experimental group required post-intervention healthcare (3 vs 10; p&#x202f;=&#x202f;0.02). CONCLUSION: Adding AI-based home exercise app to conventional treatment improve pain and may improve function and reduce healthcare utilization in rotator cuff-related shoulder pain.

Humans

Machine learning vs. traditional methods for predicting postoperative cardiac complications after non-cardiac surgery: a systematic review and Bayesian network meta-analysis.

INTRODUCTION: Accurate prediction of peri-operative cardiac complications is critical to optimise pre-operative decision-making. Traditional risk prediction scores, such as the Revised Cardiac Risk Index, show only modest discrimination. Machine learning can model complex, non-linear relationships but their predictive performance compared with traditional scores remains unclear. METHODS: We performed a systematic review and Bayesian network meta-analysis. The primary outcome was postoperative adverse cardiac events following non-cardiac surgery. Prediction models were assessed relative to the Revised Cardiac Risk Index. As many studies evaluated multiple versions of each model type, the highest performing ('best version') and lowest performing ('worst version') results were analysed. Models were ranked using the surface under the cumulative ranking curve (SUCRA). RESULTS: Thirteen studies evaluating 54 models and 927,113 patients were included. Machine learning approaches generally outperformed traditional risk scores. Automated machine learning ranked highest (SUCRA 96.6) showed the greatest improvement in the best version analysis (mean difference (MD) 0.28 (95%CrI 0.16-0.40)) and remained superior in the sensitivity analysis (MD 0.30 (95%CrI 0.14-0.45)). Gradient boosting models showed superior performance over the Revised Cardiac Risk Index across analysis (best version: MD 0.20 (95%CrI 0.14-0.26), worst version: MD 0.18 (95%CrI 0.12-0.25), SUCRA 82.4). The Gupta Perioperative Risk for Myocardial Infarction or Cardiac Arrest score outperformed the Revised Cardiac Risk Index in the best version analysis (MD 0.16 (95%CrI 0.01-0.32)). Between-study heterogeneity was low. None of the included studies externally validated their machine learning models and only six were judged to be at low risk of bias. DISCUSSION: Most machine learning models showed better discrimination than traditional risk scores, with automated machine learning and gradient boosting models ranking highest. However, study quality, calibration reporting and absence of external validation limit immediate clinical adoption. Prospective, multicentre evaluation is required before integration of these models into peri-operative practice.

Humans

Long-term microbiome and clinical effects of a microbiome-guided personalized diet versus low-FODMAP diet in irritable bowel syndrome: A 12-month follow-up randomized controlled trial.

Dietary therapy is central to irritable bowel syndrome (IBS) management, yet the long-term durability of the low-FODMAP diet (LFD), and of microbiome-guided personalization, remains unclear. We assessed the long-term clinical and gut-microbiome effects of a microbiome-guided personalized diet (PD) compared with a standard LFD in adults meeting Rome IV criteria for IBS. In this multicenter, open-label randomized controlled trial with blinded outcome assessment, participants who completed a 6-week dietary intervention (PD or LFD) were followed at 6 and 12 months without further dietary intervention. Outcomes included the IBS Severity Scoring System (IBS-SSS), IBS Quality of Life (IBS-QOL), and the Hospital Anxiety and Depression Scale (HADS); gut microbiota were profiled by 16S rRNA sequencing. Longitudinal changes were evaluated using linear mixed-effects models, responder analyses, PERMANOVA, and PERMDISP. Both diets reduced IBS-SSS at 6 weeks. PD maintained symptom improvement at 6 and 12 months (-82.0 and -78.3 points from baseline), whereas LFD benefits regressed by 12 months (+29.3 points; between-group p&#x2009;=&#x2009;0.001). At 12 months, IBS-SSS responder rates were higher with PD than LFD (62.5% vs 34.5%; absolute risk difference&#x2009;+28.0%, 95% CI 4.2-47.7; Fisher p&#x2009;=&#x2009;0.029), and IBS-QOL, HADS-anxiety, and HADS-depression showed more favourable trajectories with PD. PD was associated with sustained Shannon alpha-diversity gains (+0.488 at 6 weeks;&#x2009;+0.205 at 12 months; both p&#x2009;<&#x2009;0.01). A modest between-group beta-diversity difference at 6 months (R2&#x2009;=&#x2009;0.035; p&#x2009;=&#x2009;0.011) was not significant at 12 months. This hypothesis-generating follow-up suggests more durable benefit with PD; larger trials powered for long-term clinical and microbiome outcomes are warranted.

Humans

Manual, digital, and AI tumour-infiltrating lymphocyte scoring: a secondary analysis of the APHINITY randomised trial.

BACKGROUND: Stromal tumour-infiltrating lymphocytes (sTILs) are prognostic in early-stage HER2-positive breast cancer, but their role in the context of dual HER2 blockade remains undefined. We evaluated manual, digital, and artificial intelligence (AI)-based sTIL quantification, together with AI-derived spatial metrics, for prognostic and treatment-benefit stratification using tumour samples from the phase 3 APHINITY trial. METHODS: In the APHINITY trial, 4805 patients were randomly assigned to receive chemotherapy plus trastuzumab with pertuzumab or chemotherapy plus trastuzumab with placebo. Median follow-up was 74&#xb7;1 months (IQR 68&#xb7;3-75&#xb7;4). We analysed 4262 haematoxylin and eosin-stained images using manual assessment, an automated digital approach, AI-based lymphocyte quantification (AI percentage lymphocytes), and two AI-derived spatial features (AI-TIL and immune hotspot). Interobserver reproducibility was assessed in 262 randomly chosen tumour samples scored independently by five pathologists. Multivariable Cox models were used to assess associations between TIL levels and invasive disease-free survival (primary outcome in APHINITY), distant recurrence-free interval, and overall survival. The heterogeneity of pertuzumab benefit was evaluated using subgroup analyses, subpopulation treatment effect pattern plot analyses, and nested Cox models with treatment-by-biomarker interaction terms. FINDINGS: Manual scoring showed high interobserver reproducibility (intraclass correlation coefficient 0&#xb7;84 [95% CI 0&#xb7;79-0&#xb7;88]). Concordance between manual and automated methods was modest. AI-based scoring (AI percentage lymphocytes) reclassified 120 (11&#xb7;6%) of 1035 node-positive tumours from immune-low (by manual scoring) to immune-high; this subgroup of patients showed greater separation of 5-year invasive disease-free survival curves between pertuzumab and placebo groups compared with patients whose tumours were concordantly classified as immune-low by both manual and AI-based approaches. Higher levels of TILs were associated with improved invasive disease-free survival for all sTIL measurement approaches and spatial measurements (hazard ratios [HRs] 0&#xb7;41-0&#xb7;93). Pertuzumab was associated with improved invasive disease-free survival at higher sTIL levels across all measurement approaches (HRs 0&#xb7;36-0&#xb7;48), but was not associated with higher values of spatial measures. The largest 6-year absolute improvements with pertuzumab were observed in patients with node-positive disease whose tumours scored in the highest level of immune infiltration of manual sTIL scoring (&#x2265;70&#xb7;0%; mean absolute improvement 12&#xb7;1 percentage points [SD 2&#xb7;8]). In nested prognostic and predictive models, AI-based immune hotspot scores provided the most consistent additional information when combined with any sTIL measurement (all p<0&#xb7;010). INTERPRETATION: Standardised manual sTIL scoring was reproducible, and digital and AI-based methods showed consistent prognostic stratification and potential for treatment-benefit stratification despite only modest correlation between platforms. AI spatial metrics provided complementary information beyond sTIL density and could support more scalable immune assessment. Future studies are needed to validate these approaches in independent cohorts and to clarify their clinical utility for stratifying contemporary HER2-directed therapies. FUNDING: None.

Humans

A review into the recent advances in the world of amoebiasis.

PURPOSE OF REVIEW: Amoebiasis is a parasitic infection caused by Entamoeba histolytica , affecting 10% of the global population. It is a well recognized cause of morbidity and mortality in low-middle-income countries where it is endemic. However, with increased migration and global travel, amoebiasis is now more common in high-income countries, although diagnosis is often delayed or even missed due to lack of awareness of the latest epidemiology and optimal diagnostic testing. This review discusses the evolving prevalence, and the current international guidelines for the investigation and treatment of amoebiasis, focusing on recent advances. RECENT FINDINGS: The recent literature shows that the primary investigations for amoebiasis remain the same, though newer modalities such as artificial intelligence-powered microscopy and metagenomics have been developed recently, which aids the accuracy and speed of diagnosis. Treatment remains the same, though current research has found potential new drugs and drug targets which show promise. SUMMARY: This review reinforces the importance of early clinical suspicion, diagnosis and treatment for amoebiasis. What was once a disease only seen in endemic countries or travel-associated imported cases is now more common and must not be missed.

Humans

Metabolic engineering of Candida yeasts for biotechnological applications.

Candida yeasts represent a versatile yet underexploited platform for industrial biotechnology. These yeasts utilize a remarkably broad range of carbon sources, particularly for hydrophobic carbon sources, coupled with robust growth and diverse biosynthetic capacities, making them promising hosts for sustainable production of chemicals, fuels, and proteins. Despite these advantages, industrial deployment of Candida species has been hindered by concerns regarding opportunistic pathogenicity and the historical lack of efficient genetic manipulation tools, leading to a substantial gap between metabolic potential and practical utilization. Recent advances in functional genomics, genome editing, and systems metabolic engineering are rapidly overcoming these barriers, enabling more precise and efficient strain development. In this review, we systematically summarize recent progress in the metabolic engineering of Candida species as microbial cell factories, with particular emphasis on expanding genetic toolkits, utilizting renewable and non-conventional carbon sources, and biosynthesizing high-value compounds. In addition, we propose a biosafety-oriented classification framework to support their safe industrial deployment. Finally, we discuss current challenges and emerging opportunities, emphasizing that the synergy of synthetic biology and artificial intelligence-driven design holds the key to unlocking the biotechnological potential of Candida yeasts.

Candida