Search PubMedSearch

SEARCH · Search PubMed

Results for “AI integration”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

291 records · Page 2Linked to original sources

How AI-supported intelligent systems support infection prevention and control training in healthcare: A systematic review of educational functions and outcomes.

AIMS: Artificial intelligence (AI)-supported intelligent systems have been increasingly incorporated into infection prevention and control (IPC) education and training, primarily to support the monitoring of observable behaviors and the provision of feedback. However, existing evidence has focused largely on short-term compliance outcomes, with limited synthesis of the educational role of AI-supported intelligent systems in supporting sustained IPC competence. This systematic review examined how AI-supported intelligent systems have been designed and used to support IPC education and training, with a focus on system characteristics, educational functions, and reported outcomes. DESIGN: A systematic literature search was conducted across the PubMed/MEDLINE, Embase, Cochrane, and CINAHL databases. DATA SOURCES: A total of 18 studies met the inclusion criteria. Findings were qualitatively synthesized according to system design characteristics, educational functions, and outcome domains. REVIEW METHODS: Methodological quality was appraised using the Mixed Methods Appraisal Tool. RESULTS: Most AI-supported intelligent systems focused on hand hygiene and relied on fully automated monitoring systems to capture behaviors and provide performance feedback. Educational functions were predominantly limited to performance assessment, automated feedback, and reminders. Outcomes were mainly measured using compliance or performance metrics, whereas sustained behavioral change and decision quality were rarely assessed. CONCLUSIONS: AI-supported intelligent systems have been used primarily to reinforce short-term IPC performance and compliance. However, their current applications for supporting sustained competence over time remain limited. The findings of this review suggest that AI-supported intelligent systems may serve as maintenance-oriented educational support by extending learning beyond initial instruction through repeated practice and feedback. Future research should prioritize outcome measures that capture the durability of performance and decision-making processes to better align AI-supported intelligent systems used in IPC education and training with the educational demands of clinical practice.

Humans

Molecular Landscape and Advanced Diagnostic Technologies for BRAF Mutations in Cancer: From Quantitative PCR and ddPCR to CRISPR-Based Platforms.

BRAF mutations are key oncogenic alterations across multiple malignancies, including melanoma, thyroid carcinoma, colorectal cancer, non-small cell lung cancer, glioma, and hairy cell leukemia. The most prevalent variant, BRAF-V600E, induces constitutive activation of the MAPK signaling pathway, promoting tumor progression and influencing therapeutic responsiveness. Accurate detection of BRAF alterations is therefore essential for molecular classification, prognostic assessment, treatment selection, and resistance surveillance. This review summarizes the molecular heterogeneity of BRAF mutations and critically evaluates current diagnostic methodologies. Conventional approaches such as allele-specific PCR and Sanger sequencing are compared with advanced quantitative platforms, including high-resolution melting analysis, droplet digital PCR, and next-generation sequencing, with emphasis on analytical sensitivity, mutation coverage, and clinical applicability. Emerging technologies such as CRISPR-based assays, rolling circle amplification systems, and nanoparticle-based biosensors and point-of-care diagnostic platforms are also discussed for their potential to enhance ultra-sensitive detection, particularly in liquid biopsy settings. These emerging tools are highlighted for their potential to enable ultra-sensitive, rapid, and decentralized mutation detection, particularly in liquid biopsy settings. Key challenges, including intratumoral heterogeneity, low allele-frequency variants, FFPE-associated artifacts, and clonal evolution under therapeutic pressure, are examined within a translational framework. In addition, we examine critical barriers to clinical implementation, including standardization, cost, and global accessibility of molecular diagnostics, and outline potential solutions through scalable technologies and decentralized testing strategies. We propose that optimal BRAF testing requires a mutation subclass-informed and clinically integrated strategy combining comprehensive baseline profiling with longitudinal molecular monitoring. Future diagnostic paradigms will likely integrate multi-omics data and artificial intelligence (AI)-assisted interpretation to refine precision oncology implementation. Looking forward, we propose that optimal BRAF testing will require integration of multi-omics profiling with AI-assisted interpretation, enabling automated variant classification, real-time clinical decision support, and improved prediction of therapeutic response and resistance.

Humans

Enhanced fracture detection on radiographs with AI assistance for clinicians: a systematic review and meta-analysis.

BACKGROUND: Emergency radiographic interpretation for fractures is prone to missed or misdiagnoses. Artificial intelligence (AI) is expected to become a powerful tool to assist clinicians in fracture detection. PURPOSE: A systematic review and meta-analysis was performed to assess whether AI improves clinicians' ability to detect fractures on radiographs. MATERIALS AND METHODS: A literature search was conducted in PubMed, Web of Science, and Cochrane Library for studies published between January 1, 2010, and October 10, 2025. A meta-analysis of diagnostic accuracy studies was performed using a Summary Receiver Operating Characteristic (SROC) curve. The quality of included studies was assessed using the Quality Assessment of Diagnostic Accuracy Studies 2 (QUADAS-2) tool. Subgroup analysis and meta-regression were conducted to explore potential sources of heterogeneity. RESULTS: A total of 26 studies were included . The pooled sensitivity of clinicians increased from 77% (95% CI: 72-81) to 87% (95% CI: 83-90) with AI assistance, while the pooled specificity improved from 88% (95% CI: 85-90) to 92% (95% CI: 89-94). The corresponding AUC values were 0.90 (95% CI: 0.87-0.92) before and 0.95 (95% CI: 0.93-0.97) after AI assistance. Eight studies were rated as high risk of bias. Subgroup analysis and meta-regression identified potential sources of heterogeneity, including fracture location, AI model type, high risk of bias, and reference standards. CONCLUSION: AI assistance significantly improves clinicians' diagnostic performance in detecting fractures on radiographs for extremity and trunk fractures.

Humans

Experimental validation of an AI-driven digital healthcare platform for oral health behavior and plaque assessment among vietnamese children.

BACKGROUND: Oral health among children in developing countries, including Vietnam, remains a significant public health concern. Innovative approaches leveraging artificial intelligence AI-based digital health platforms may offer effective strategies for managing dental plaque and promoting better oral hygiene behaviors among school-aged children. This study aimed to evaluate the effectiveness of an AI-driven oral healthcare platform (Denti-i Vietnam) in improving oral hygiene and behavioral outcomes among Vietnamese primary school students. METHODS: A total of 204 primary school students aged 8-10&#xa0;years in Hanoi, Vietnam, participated in this experimental study. Participants were randomly assigned to an intervention group (n&#xa0;=&#xa0;107), which used the AI-driven oral healthcare platform, and a comparison group (n&#xa0;=&#xa0;97), which received traditional oral health education via pamphlets. Oral health behaviors, dental plaque levels (Simplified Oral Hygiene Index; OHI-S), and caries indices (dft/DMFT) were assessed at baseline and after the intervention period. RESULTS: The intervention group demonstrated a significant reduction in the OHI-S score compared to baseline (2.49&#xa0;&#xb1;&#xa0;0.60 to 1.70&#xa0;&#xb1;&#xa0;0.76, p&#xa0;<&#xa0;0.001), particularly in the debris component, indicating enhanced plaque control. Notable improvements were also observed in oral hygiene behaviors, including increased frequency of toothbrushing before and after breakfast (p&#xa0;<&#xa0;0.01) and more frequent parental assistance during brushing (p&#xa0;=&#xa0;0.03). Furthermore, parental awareness of dental caries significantly increased in the intervention group (p&#xa0;=&#xa0;0.001). CONCLUSIONS: The AI-driven oral healthcare platform significantly improved both oral hygiene behaviors and plaque control among Vietnamese primary school children. These findings suggest that AI-driven digital health tools can serve as practical and scalable solutions for promoting oral health in developing countries.

Humans

Beyond predictive performance: A systematic review and critical methodological appraisal of AI/ML and conventional modelling strategies in breast, colorectal, and pancreatic Cancer.

BACKGROUND: Predictive modelling for cancer risk, treatment-related complications, and survival is central to precision oncology. Conventional logistic regression (LR) and Cox proportional hazards (CoxPH) regression remain widely used but are limited when modelling nonlinear interactions, high-dimensional imaging features, and multimodal clinical-metabolic predictors. Artificial intelligence (AI) and machine learning (ML) methods offer expanded capability through automated feature extraction, ensemble learning, and flexible survival modelling, but the evidence on when AI/ML adds value over conventional models across cancer sites and predictive tasks remains fragmented. OBJECTIVE: To systematically evaluate the methodological performance, validation strategies, and translational limitations of AI/ML models compared with conventional statistical models in published predictive-modelling studies for breast, colorectal, or pancreatic cancer. METHODS: PubMed, Scopus, and Web of Science were searched for studies published between January 2019 and March 2025. Two reviewers independently conducted title-and-abstract screening, full-text eligibility assessment, and PROBAST risk-of-bias assessment. Sixty-five studies (n&#xa0;=&#xa0;907,567 participants) were narratively synthesised by cancer site, predictive task, model family, comparator, validation strategy, predictor modality, and calibration or explainability reporting. RESULTS: The 65 studies comprised breast cancer (n&#xa0;=&#xa0;35), colorectal cancer (n&#xa0;=&#xa0;21), and pancreatic cancer (n&#xa0;=&#xa0;9). AI/ML superiority over LR and CoxPH was task- and data-dependent. CNN- and U-Net-based models predominated in imaging and body-composition tasks, tree-based ensembles consistently outperformed LR for tabular perioperative complication prediction, and CoxPH remained competitive, and in the largest pancreatic risk study, superior to XGBoost (C-index 0.802 vs 0.723) in well-structured datasets. PROBAST analysis-domain risk was moderate in 54 of 65 studies (83%), driven by limited external validation, sparse calibration reporting (11/65), and few decision-curve analyses (7/65). CONCLUSION: AI/ML adds the most methodological value in imaging-derived feature extraction and nonlinear perioperative prediction, while conventional regression remains preferable in large, structured datasets with linear predictors. Clinical translation requires standardised body-composition definitions, external validation, calibration assessment, decision-curve analysis, and explainability, in line with TRIPOD+AI and CLAIM standards.

Humans

Effectiveness of an AI-based home exercise app for rehabilitation of rotator cuff-related shoulder pain: A randomized controlled trial.

BACKGROUND: Rotator cuff-related shoulder pain contributes to disability and healthcare use. Although therapeutic exercise is first-line treatment, limited supervision and adherence may reduce its effectiveness; digital rehabilitation with real-time feedback may address these limitations. OBJECTIVES: To evaluate the effectiveness of adding a digital rehabilitation program to standard physiotherapy on pain, function, fear-avoidance beliefs, and healthcare utilization. DESIGN: Single-center, assessor-blinded, randomized controlled trial with two parallel groups. METHOD: Forty-six adults (mean age 59 years) with rotator cuff-related shoulder pain were randomized to 12 weeks of conventional physiotherapy or physiotherapy plus an AI-based digital rehabilitation program using computer vision for real-time feedback and performance monitoring. Outcomes were assessed at baseline and at 2, 4, and 12 weeks. Pain intensity (NPRS) was primary outcome; secondary outcomes included upper limb function (QuickDASH), fear-avoidance beliefs (FABQ), and post-intervention healthcare utilization. Analyses followed an intention-to-treat approach. RESULTS: Pain reduction exceeded the MCID (1.3) at 4 and 12 weeks. Between-group differences favoured the intervention at Weeks 2 and 4 (MD -0.7; 95% CI -1.13 to -0.14 and MD -1.01; 95% CI -1.8 to -0.2, respectively). Upper limb function improved more at Week 4 (MD -7.3; 95% CI -12.3 to -2.2). FABQ scores decreased more at Week 12 (MD -7.6; 95% CI -14 to -0.5). Fewer participants in the experimental group required post-intervention healthcare (3 vs 10; p&#x202f;=&#x202f;0.02). CONCLUSION: Adding AI-based home exercise app to conventional treatment improve pain and may improve function and reduce healthcare utilization in rotator cuff-related shoulder pain.

Humans

Applications of quantum AI in brain disorder diagnosis: A systematic review.

BACKGROUND AND OBJECTIVE: Brain disorder diagnosis and prediction remain challenging because neuroimaging, electrophysiological, behavioral, and multimodal data are high-dimensional, noisy, heterogeneous, and limited by small clinical cohorts. This systematic review synthesised applications of quantum artificial intelligence (QAI) for brain disorder diagnosis, prediction, detection, and monitoring. METHODS: Following PRISMA guidelines, studies published from 2016 to 13 January 2026 were retrieved from Scopus, Web of Science, and IEEE Xplore. After screening, 36 studies met the eligibility criteria and were qualitatively analysed according to disorder category, data modality, QAI method, implementation setting, validation strategy, and performance. RESULTS: At the broader disease-group level, neurodegenerative disorders were the most frequently investigated, followed by mental health and psychiatric disorders. At the individual level, Parkinson's disease and schizophrenia were the leading applications, followed by depression, anxiety, Alzheimer's disease, and stress-related tasks. MRI-based modalities were the most frequently used data source, followed by multimodal data and EEG. Methodologically, primary QAI approaches were dominated by quantum neural and QDL architectures, followed by quantum-inspired optimization or feature-selection methods and quantum-kernel/conventional QML classifiers. Qiskit/IBM Quantum and PennyLane were the most frequently reported quantum software frameworks. However, most studies relied on simulators, classical quantum-inspired implementations, or unclear implementation settings, with limited real-hardware evaluation. CONCLUSIONS: QAI shows emerging potential for brain disorder analysis, particularly through hybrid quantum-classical learning, quantum neural architectures, quantum-kernel methods, and quantum-inspired optimization. Nevertheless, current evidence remains preliminary and requires larger datasets, subject-level and external validation, fair classical benchmarking, noise-resilient circuits, real quantum hardware evaluation, explainability, and clinical validation.

Humans

Probiotic-derived extracellular vesicles as food-based nanocarriers: Mechanisms, functional applications, and future perspectives in food systems.

Probiotic-derived extracellular vesicles (PDEVs) are a promising type of postbiotic nanoparticle derived by fermentation of probiotics, and have gained growing interest as a potential application in food science and nutrition. These are lipid bilayer vesicles of nanoscale, which are naturally released by probiotic cells and contain a wide variety of bioactive molecules, such as proteins, nucleic acids, and metabolites. Moreover, PDEVs are highly stable, biocompatible, and can be easily engineered to have surfaces with high functionality, which makes them good candidates in functional engineering. In contrast to traditional live probiotics, PDEVs overcome the difficulties of preserving microbial viability during processing and storage, thus providing superior safety, stability, and predictable biological performance. This is a systematic review of the various functions of PDEVs in food systems. We conclude on the processes through which PDEVs control intestinal barrier integrity, alter gut microbiota composition, and alter host immune responses, and their potential to enhance gut health when added to functional foods. In addition to their health-promoting effects, PDEVs have shown significant potential as natural antimicrobial agents to preserve food and as effective nanocarriers of hydrophobic bioactive compounds, including fucoxanthin, to improve their stability, bioavailability, and targeted delivery. Moreover, PDEVs can be used as new regulators of microbial fermentation. However, it should be noted that a lot of the evidence that is available is still preliminary and the effectiveness of these applications in real food-processing and storage conditions has not been fully proven. Although they have potential, there are a number of challenges that still hinder the widespread use of PDEVs in the food industry. These involve the creation of scalable and cost-effective production processes, batch-to-batch consistency, vesicle stability in a variety of food matrices, and regulatory and safety considerations. Other emerging engineering approaches, such as surface functionalization and cargo loading, are also discussed in this review and could further increase the specificity, functionality, and application versatility of PDEVs in food systems. Moving forward, the incorporation of PDEVs into the next generation functional foods, novel food preservation methods, and customized nutrition plans should be prioritized in future studies. Further developments in these fields can make PDEVs useful platforms at the interface of food microbiology, nanotechnology, and human health.

Probiotics

Artificial intelligence-derived myocardial fibrosis on cardiac magnetic resonance for prognosis in cardiomyopathy: A systematic review of a sparse evidence base.

BACKGROUND: Myocardial fibrosis on cardiovascular magnetic resonance (CMR), assessed by late gadolinium enhancement (LGE) and parametric mapping, is an established predictor of adverse events in cardiomyopathy. We assessed whether artificial intelligence (AI) quantification of fibrosis adds independent prognostic value. METHODS: We searched six databases, a clinical-trials register, and a preprint server from inception to 13 June 2026. Eligible studies used AI to generate a fibrosis marker in adults with ischemic or nonischemic cardiomyopathy, with covariate-adjusted outcomes over &#x2265;12 months. Risk of bias was assessed using PROBAST, PROBAST+AI, and QUIPS. Fewer than three comparable studies precluded meta-analysis; certainty was rated using GRADE. RESULTS: Of 448 records (381 after de-duplication), 18 full texts were reviewed and two included, one peer-reviewed and one preprint. In an ischemic-cardiomyopathy registry (Ghanbari et al.; n = 216 analytic, 26 events), AI-derived dense LGE scar predicted arrhythmic events (univariable hazard ratio [HR] 2.35, 95% CI 1.33-4.15), and AI-derived but not manual scar improved discrimination beyond guideline criteria (area under the curve 0.63 to 0.68; p = 0.02). In a nonischemic dilated-cardiomyopathy preprint (Kim et al.; n = 347, 119 events), automated extracellular volume &#x2265;30% predicted cardiovascular death or heart-failure hospitalization (adjusted HR 2.00, 95% CI 1.32-3.03). Both were at high risk of bias, with data-derived thresholds and no external validation. CONCLUSIONS: Across only two studies, AI-derived fibrosis was independently associated with adverse cardiovascular events, but its added value over manual quantification remains unproven. Certainty was very low. The evidence base is sparse and not yet ready for clinical use.

Humans

Diagnostic performance of machine learning models for malignant and non-malignant pleural effusion: Systematic review and meta-analysis.

BACKGROUND: Accurately distinguishing malignant pleural effusion (MPE) from non-malignant pleural effusion is clinically important, but the generalisability and methodological quality of machine-learning (ML) models remain uncertain. METHODS: We searched eight databases to 23 April 2026. Diagnostic performance was pooled using random-effects and Reitsma bivariate models, and study quality was assessed using PROBAST+AI. RESULTS: Forty-two studies were included; 17 contributed to the AUC meta-analysis and 14 to the bivariate analysis. The pooled AUC was 0.90 (95&#xa0;% CI 0.85-0.94; 95&#xa0;% prediction interval 0.62-0.98), with sensitivity of 0.80 (95&#xa0;% CI 0.77-0.83) and specificity of 0.87 (95&#xa0;% CI 0.79-0.92). Only nine studies reported external, temporal or independent validation. Externally validated studies had a lower pooled AUC than studies without external validation (0.83 vs 0.92), with lower specificity observed in the two externally validated studies contributing sensitivity and specificity data. All 42 development assessments had high overall quality concerns, and all 42 model evaluations were judged at high risk of bias. CONCLUSIONS: ML models showed good apparent accuracy for distinguishing MPE from non-MPE, but the evidence was limited by substantial heterogeneity, high risk of bias and scarce external validation. The pooled estimates reflect the average performance of different selected models rather than the expected accuracy of a single clinical test. ML models should be regarded as adjuncts to existing diagnostic pathways until they are confirmed by rigorous multicentre prospective external validation and clinical-impact studies.

Humans

Combining neuromelanin-sensitive MRI and quantitative susceptibility mapping for enhanced diagnosis and differentiation of parkinson's disease: A systematic review.

BACKGROUND: Loss of dopaminergic neurones and iron deposition in the substantia nigra pars compacta (SNpc) are two major pathological hallmarks of Parkinson's disease (PD). Such changes can be visualised by advanced techniques including neuromelanin-sensitive MRI (NM-MRI) and quantitative susceptibility mapping (QSM). This systematic review investigates the diagnostic performance and methodological development of the integrated use of NM-MRI and QSM in PD. METHODS: The systematic search was performed in four databases (Scopus, PubMed, ScienceDirect, and Web of Science) according to the PRISMA 2020 guidelines until July 2026. Bias was assessed using QUADAS-2 and certainty of evidence was assessed using GRADE. RESULTS: Seventeen studies with 2228 participants were included. Combined NM-MRI and QSM consistently showed reduced neuromelanin volume/contrast and increased iron deposition in the SNpc of PD patients compared to healthy controls. Multimodal integration yielded a significant improvement in diagnostic accuracy (AUC values 0.86-0.99), and was able to successfully differentiate PD. Recent methodological advances included simultaneous acquisition sequences (e.g. MTC-GRE, STAGE, setMag) and AI-driven automated segmentation, which led to significantly reduced scan times and improved reproducibility. CONCLUSION: The combination of NM-MRI and QSM has a synergistic effect and provides powerful complementary biomarkers for the diagnosis and differential diagnosis of PD.

Humans

The impact of artificial intelligence on critical thinking and clinical reasoning in health professions education: A systematic review and meta-analysis.

BACKGROUND: Critical thinking and clinical reasoning underpin healthcare professionals' ability to navigate uncertainties and deliver safe and effective care. With artificial intelligence (AI) advancement and growing adoption, AI-based educational tools are increasingly used to support these cognitive competencies' development. OBJECTIVE: To synthesize randomised and controlled clinical trials on AI-based educational tools in health professions education and examine their effects on critical thinking and clinical reasoning among health professions students. METHODS: Six electronic databases were searched from January 1, 2014 to July 28, 2025 was reviewed: PubMed, Cochrane Central Register of Controlled Trials, CINAHL, Scopus, Embase and Web of Science. Two independent reviewers performed data extraction and quality assessment using standardized JBI checklists. The GRADE approach was used to assess the certainty of evidence. Studies were pooled via random-effects meta-analyses or narrative syntheses. RESULTS: Fourteen randomised controlled trials and seven controlled clinical trials were included (n&#xa0;=&#xa0;21). Meta-analyses revealed small to medium effect sizes for the surrogate clinical reasoning outcomes of performance-based assessment scores (SMD 0.68; 95% CI [0.38, 0.98], p-value&#xa0;=&#xa0;0.00; I2&#xa0;=&#xa0;38%) and knowledge test scores (SMD 0.39; 95% CI [0.09, 0.69], p-value&#xa0;=&#xa0;0.01; I2&#xa0;=&#xa0;79%). Critical thinking and clinical reasoning skills and dispositions were narratively synthesized, with majority of included studies favouring AI-based interventions but the evidence had low to very low certainty. CONCLUSION: AI-based educational interventions may improve critical thinking and clinical reasoning among health profession students, but the evidence is very uncertain. This review offers preliminary insights but does not allow identification of optimal interventions or discipline-specific recommendations due to small sample sizes and substantial intervention heterogeneity. Further research is required to draw definitive conclusions. PROTOCOL REGISTRATION: CRD42025634074.

Humans

Phase 1 Study Evaluating Gefurulimab Pharmacokinetics and Safety Following Delivery Via Autoinjector or Prefilled Syringe With Needle Safety Device in Healthy Adults.

PURPOSE: Gefurulimab, a novel dual-binding nanobody targeting complement component 5 (C5), is in clinical development for anti-acetylcholine receptor antibody-positive generalized myasthenia gravis. Gefurulimab has a low molecular weight, enabling subcutaneous (SC) self-administration by autoinjector (AI) or prefilled syringe with needle safety device (PFS-SD). We compared gefurulimab pharmacokinetic (PK) exposure and safety in healthy adults following a single SC dose administered by AI versus PFS-SD. METHODS: In this phase 1, open-label, randomized, parallel-group study (NCT06208488), healthy participants aged 18 to 65 years were stratified by weight and randomized equally to 1 of 6 combination groups of device and injection site (abdomen/thigh/upper arm). Participants received a single SC dose of gefurulimab on day 1 and were assessed throughout the 92-day evaluation period. Primary endpoints were PK parameters for each device: maximum observed concentration (Cmax) and area under the serum concentration-time curve (AUCinf, AUClast). PK across injection sites, pharmacodynamics, safety, immunogenicity, and device performance were also assessed. FINDINGS: Overall, 175 participants were randomized: AI (n = 87), PFS-SD (n = 88). Geometric least squares mean ratios (90% CI) comparing AI/PFS-SD for Cmax, AUCinf, and AUClast were 97.6% (94.5-100.8), 99.6% (96.1-103.3), and 98.8% (95.2&#x2012;102.6), respectively. Secondary analyses found no meaningful differences in PK parameters across injection sites. Serum-free C5 concentrations over time, treatment-emergent adverse event (TEAE) profiles, and antidrug antibody responses were similar between cohorts. Most TEAEs were mild; none led to study discontinuation. IMPLICATIONS: SC administration of gefurulimab by AI and PFS-SD was well tolerated with comparable exposure, meeting bioequivalence criteria.

Humans

The future of pediatric vesicoureteral reflux management.

BACKGROUND AND OBJECTIVE: Vesicoureteral reflux (VUR) is a common condition in pediatric urology, yet important uncertainties persist regarding risk stratification, imaging strategies, and prevention of long-term renal damage. Emerging technologies may help address these challenges. This review provides a forward-looking overview of recent advances in artificial intelligence (AI) and immunomodulation that may influence future management of pediatric VUR. METHODS: A forward-looking literature review was performed using the PubMed database (January 2000-March 2025), focusing on studies addressing AI, immunomodulation, or vaccination in the context of VUR and urinary tract infections. Criteria of inclusion were the relevance to pediatric VUR, the novelty of the proposed concept, the potential clinical implications and, for the AI literature, the existence of a clinical evaluation of the algorithm on a dataset from patients. KEY FINDINGS AND LIMITATIONS: AI-based models show promising performance in supporting clinical decision-making, including prediction of the need for voiding cystourethrography, automated grading of VUR, estimation of recurrent urinary tract infection risk and prediction of chemoprophylaxis. These tools may facilitate more individualized diagnostic and therapeutic strategies, although current evidence is largely retrospective and requires prospective validation. Immunization and immunomodulatory approaches aim to reduce infection burden and modulate inflammatory pathways associated with renal scarring. While early experimental and adult clinical data are encouraging, pediatric-specific evidence remains limited, and clinical applicability in children with VUR is not yet established. CONCLUSION: Artificial intelligence and immunologically targeted strategies represent complementary, emerging approaches that may contribute to more personalized management of pediatric VUR. At present, both should be regarded as exploratory tools whose clinical impact will depend on further validation and appropriately designed pediatric studies.

Humans

Artificial intelligence for anticancer drug discovery from natural products of macroalgae and sponges: A systematic review.

Marine natural products (MNPs) from macroalgae and marine sponges have inspired clinically important anticancer agents, including the cytarabine pharmacophore and the eribulin scaffold, while cyanobacterial dolastatin chemistry supplies the auristatin payloads of several marine-inspired antibody-drug conjugates (ADCs) such as brentuximab vedotin. Artificial intelligence (AI) methods, encompassing both classical machine learning (ML) with hand-engineered features and modern deep learning (DL) with many-layered neural networks, are increasingly supporting key decisions in natural-product anticancer drug discovery, including bioactivity prediction, target identification, absorption, distribution, metabolism, excretion and toxicity (ADMET) filtering, generative analogue design, and the selection of preclinical candidates. DL architectures relevant to this field include graph neural networks, transformer-based molecular generators, diffusion models for protein-ligand docking, and convolutional networks for mass spectrometry, while classical ML contributes interpretable fingerprint-based bioactivity models and molecular networking for dereplication. This review follows a systematic literature review methodology to organize the landscape of AI methods now applied to MNP anticancer discovery, distinguishing ML and DL approaches where relevant, situating them within the chemical context of macroalgal and sponge-derived oncology leads, and critically examining published case studies, including validation level (computational, in vitro, in vivo, clinical). The principal bottleneck for medical translation has shifted partly from algorithmic capability toward data infrastructure and experimental validation. Sparse, heterogeneous, and taxonomically biased bioactivity records limit what current models can learn and reduce the reliability of AI-prioritized candidates entering the preclinical pipeline. A roadmap is proposed that prioritizes open MNP-specific benchmarks, symbiont-aware modeling, and active learning loops with synthesizability and ADMET constraints. These AI workflows may accelerate the prioritization of marine-derived anticancer leads and support earlier, more evidence-based translational decisions in oncology drug development.

Biological Products

Can ChatGPT Replace Human Clinical Coders? A Comparative Study in Otology Billing.

OBJECTIVE: Evaluate the utility of the large language model (LLM), ChatGPT, for the analysis of operative notes and the generation of Current Procedural Terminology (CPT) codes in comparison to human clinical coders. STUDY DESIGN: CPT billing codes assigned by ChatGPT were compared to existing billing data. Otology practice within a tertiary academic center. METHODS: About 191 operative notes from a single surgeon (9/2022-10/2023) were analyzed. ChatGPT-3.5 and 4 models were prompted for CPT codes based on operative notes. Assessment included determining exact and partial match rates, sensitivity and specificity for targeted procedures, and work Relative Value Units (wRVU) differences between ChatGPT-generated and human-assigned codes. RESULTS: ChatGPT-3.5 achieved exact matches in 22% of cases and partial matches in 32%, while ChatGPT-4 achieved 14% exact and 33% partial matches. When cochlear implantation (CI) was excluded, performance dropped significantly. For CI, ChatGPT-3.5 demonstrated a sensitivity of 94% and specificity of 90%, while ChatGPT-4 showed a sensitivity of 96% and specificity of 92%. In contrast, performance on cartilage grafting was poor, with sensitivities of 4.2% for ChatGPT-3.5 and 0% for ChatGPT-4. ChatGPT-3.5 and 4 showed moderate CPT code matching accuracy among themselves, with slight agreement to human coders. Both models tended to underbill for wRVUs compared to human coders, with significant differences in the values generated. CONCLUSION: This study assessed ChatGPT's effectiveness in automating CPT code assignment for otologic surgeries. While the models achieved high sensitivity values for assigning codes related to cochlear implantation, both models struggled with complex cases, failed to apply modifiers, and often assigned fewer wRVUs. The findings highlight ChatGPT's potential in medical billing but indicate a need for further refinement.

Humans

Data-centric, robust, and explainable multimodal deep learning for clinical decision support: A systematic review.

PURPOSE: Multimodal deep learning is increasingly proposed for clinical decision support (CDS) under a "data-centric" framing that prioritizes label quality, missing-modality robustness, distribution shift, calibration, and explainability. Prior reviews have examined multimodal medical AI, CDS, and data-centric methods separately, but none address their intersection. We mapped the modalities, fusion strategies, and data-centric and explainability techniques used in this recent literature, quantified how often each is implemented rather than merely mentioned, assessed deployment-relevant evidence (external validation, clinical-outcome measurement, equity), and formally appraised study-level risk of bias. METHODS: Following the PRISMA 2020 statement (PROSPERO CRD420261427815; registered retrospectively), we screened 150 records and included primary, clinical, multimodal studies that applied machine or deep learning to a decision-support task and reported at least one quantitative result. Two reviewers screened and extracted data with consensus adjudication. Each study was coded against pre-specified operational definitions, separating implemented or empirically evaluated techniques from those only mentioned. Study-level risk of bias was assessed with PROBAST + AI. Synthesis was narrative. RESULTS: Thirty-one studies met inclusion; 30 (97%) were published between 2024 and 2026, with a median of three modalities (range 2-6), most commonly structured EHR (71%) and imaging (39%). Data-centric techniques were frequently reported (74-84% across label-noise, distribution-shift, calibration, missing-modality and class-imbalance handling; equity 61%). However, external validation was reported in only 4/31 studies (13%), a clinical or provider outcome in 3/31 (10%), and no study reported routine deployment. Overall risk of bias was high in 27/31 studies (87%), driven by the analysis domain. CONCLUSION: Within this recent, self-selected slice of the field, technical robustness and explainability techniques are widely reported but rarely validated out-of-distribution or against clinical outcomes, and the underlying evidence is at high risk of bias. Progress requires external multi-site validation, clinical-outcome measurement, formal bias appraisal, and adherence to AI reporting standards (e.g., TRIPOD + AI) before deployment can be justified.

Deep Learning

Artificial intelligence-assisted detection and optical differentiation of colorectal lesions in Lynch syndrome surveillance (CADLY2): a multicentre, open-label, randomised controlled superiority trial.

BACKGROUND: Artificial intelligence (AI)-based computer-aided detection (CADe) systems improve adenoma detection in average-risk colorectal cancer screening. Meanwhile, evidence in Lynch syndrome surveillance is sparse and inconsistent. We assessed the effect of CADe on adenoma detection during Lynch syndrome surveillance. Computer-aided optical diagnosis (CADx) performance for optical differentiation of colorectal lesions was evaluated as a secondary aim. METHODS: CADLY2 was an international, multicentre, open-label, randomised controlled superiority trial at nine specialised hereditary cancer surveillance centres in Belgium, Germany, the Netherlands, and Spain. Adults aged 18 years or older with genetically confirmed Lynch syndrome scheduled for surveillance colonoscopy were randomly assigned (1:1) to high-definition white-light (HD-WL) colonoscopy alone or to HD-WL colonoscopy with computer-aided assistance from CAD EYE (Fujifilm, Tokyo, Japan). CAD EYE was used for CADe during withdrawal and for CADx after lesion detection. Randomisation was done centrally through a secure web-based system using Pocock's minimisation algorithm with a stochastic component and was stratified by centre, sex, previous colorectal cancer, underlying pathogenic variant, and interval since previous colonoscopy. Allocation concealment was ensured through the centralised web-based system. Patients were masked to group allocation until the start of withdrawal in procedures with mild sedation, or until completion of the procedure in procedures with propofol-based sedation. Endoscopists were not masked. The primary outcome was adenoma detection rate, defined as the proportion of patients with at least one histopathologically confirmed adenoma, analysed in the full analysis set (defined as all randomly allocated patients with available data for the primary outcome). The diagnostic performance of the CADx system was evaluated as a secondary outcome. The safety analysis set comprised all randomly allocated patients who underwent a study colonoscopy. This study is registered with the German Clinical Trials Register, DRKS00030695, and is completed. FINDINGS: Between May 9, 2023, and Oct 30, 2025, 757 patients were randomly allocated to HD-WL colonoscopy (377 patients) or to AI-assisted colonoscopy (380 patients); 733 patients were included in the full analysis set (369 HD-WL and 364 AI-assisted). The median age was 49 years (IQR 38-59) in the HD-WL group and 50 years (38-59) in the AI-assisted group; 213 (58%) were female and 156 (42%) male in the HD-WL group, and 207 (57%) were female and 157 (43%) male in the AI-assisted group. The adenoma detection rate was 30&#xb7;9% (114 of 369 patients) with HD-WL versus 33&#xb7;8% (123 of 364 patients) with CADe assistance (odds ratio 1&#xb7;14 [95% CI 0&#xb7;83-1&#xb7;57], p=0&#xb7;41). For CADx differentiation of neoplastic versus non-neoplastic lesions in the paired lesion-level analysis, with histopathology as the reference standard and sessile serrated lesions and traditional serrated adenomas classified as non-neoplastic, CADx sensitivity was 85&#xb7;9% (95% CI 82&#xb7;0-89&#xb7;1) and specificity was 91&#xb7;4% (89&#xb7;4-93&#xb7;0). Three adverse events occurred in the AI-assisted group: two mild post-polypectomy bleedings and one serious pulmonary embolism or deep venous thrombosis unrelated to the procedure. No adverse events occurred in the HD-WL group. INTERPRETATION: CADe-assisted colonoscopy did not show the absolute improvement in adenoma detection rate that was assumed in the prespecified sample-size calculation. CADx did not clearly improve lesion differentiation beyond expert optical diagnosis in expert Lynch syndrome surveillance settings. FUNDING: Third-party research funding of the National Center for Hereditary Tumor Syndromes, University Hospital Bonn.

Humans