Search PubMedSearch

SEARCH · Search PubMed

Results for “learning curve”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

233 records · Page 3Linked to original sources

Effects of extended problem-based learning interventions on undergraduate nursing education: A systematic review.

OBJECTIVE: Exploring the effects of long-term PBL (problem-based learning) intervention on undergraduate nursing students. METHODS: The article retrieved literature from CINAHL Complete, Academic Search Complete, Web of Science, PubMed, EMBASE, OVID, and Cochrane Library up to January 2025. Studies had to meet all of these criteria: (1) They used a randomized controlled trial (RCT) and quasi-experimental design. (2) The PBL pedagogy intervention lasted 4 weeks or longer. (3) The participants were undergraduate nursing students. (4) They reported primary outcomes. These included critical thinking, problem-solving skills and self-directed learning. Two researchers screened articles, extracted data, and assessed quality independently using blinding. They used Cochrane ROB2 for RCTs and ROBINS-I for quasi-experimental studies to judge bias risk. Meta-analysis was performed using RevMan 5.4 software. For continuous variables, standardized mean difference (SMD) and 95% confidence interval were calculated. Heterogeneity was assessed by I2 statistic. When I2 > 50%, sensitivity analysis was conducted. The source of heterogeneity was explored by excluding studies one by one. The primary outcomes included standardized critical thinking, problem-solving, and self-directed learning assessment results. RESULTS: A total of 11 randomized controlled trials and quasi-experimental studies were retrieved and included for meta-analysis. The experimental group significantly outperformed the control group in critical thinking, problem-solving, and self-directed learning, with differences being statistically significant (P ≤ 0.05). However, high heterogeneity was observed. After sensitivity analysis, the heterogeneity was reduced and the results remained statistically significant, indicating that the findings were not solely dependent on the excluded studies.

Problem-Based Learning

Comparative effectiveness of game-based learning modalities in nursing and medical education: a systematic review and Bayesian network meta-analysis.

BACKGROUND: Game-based learning (GBL) is increasingly used in healthcare education, but educators must choose among diverse modalities (e.g., quiz platforms, apps, serious games and metaverse environments). Comparative evidence on which modalities perform best across learning domains (knowledge, attitudes, and practice) remains limited. AIM: To compare the effects of distinct GBL modalities on knowledge, attitudes, and practice outcomes in nursing and medical education and to explore whether comparative effects differ by learner group (pre-licensure students and in-service professionals). DESIGN: PRISMA-NMA-aligned systematic review and Bayesian network meta-analysis. METHODS: We searched eight databases and trial registries through September 2, 2024, for randomized controlled trials comparing GBL with traditional teaching (TT). Outcomes were transformed to a 0-100 scale and analysed as change from baseline in Bayesian consistency models; random-effects models were selected using deviance information criterion (DIC). Risk of bias was assessed using RoB 2. We report mean differences (MDs) with 95% credible intervals (CrIs) versus TT, ranking probabilities, and subgroup NMAs by learner group. RESULTS: Thirty-one RCTs (n = 3439) were included; 15 contributed complete data to the network. Risk of bias was low in 15 trials and raised some concerns in 16. The network was modest for knowledge (11 trials) and sparse for attitudes (3) and practice (4). Compared with TT, metaverse-based learning showed improved attitudes (MD 15; 95% CrI 12 to 18), based on a single trial. For knowledge and practice, Kahoot-based quizzes (MD 9.1; 95% CrI -8.9 to 27) and app-based learning (MD 4.6; 95% CrI -4.4 to 14) had the highest estimated mean improvements, but credible intervals were wide and included the null for most comparisons. Subgroup rankings differed by learner group, but several comparisons were imprecise and uncertainty was substantial, particularly in sparse networks. CONCLUSIONS: GBL modalities may improve learning outcomes compared with TT, but relative effects appear domain-specific and the certainty of rankings is limited by sparse evidence and imprecision. Future trials should prioritise head-to-head comparisons, robust outcome measurement, and longer-term retention and transfer outcomes in both student and in-service populations.

Humans

Does Inquiry-Based Learning Improve Students' Critical Thinking? A Meta-Analysis Accounting for Control Group Variations.

BACKGROUND: Inquiry learning is widely recognized, through empirical studies, as an appropriate instruction in enhancing students' critical thinking, yet the results were varied across context. The previous meta-analysis did not include the control group variations as a potential moderator and the studies subject domain was limited only to science subjects. Consequently, it is difficult to generalize the effectiveness of IBL in enhancing critical thinking. This meta-analysis aims to investigate whether inquiry learning is effective in improving the students' critical thinking skills and examine the moderating roles of each study characteristic. Methods The literature search applying the PRISMA protocol 2020 was conducted by utilizing SCOPUS, ERIC, and DOAJ databases. A total of 57 studies from 51 articles, published from 2015 to 2025, were synthesized using a random-effects model with standardized mean difference (SMD). RESULTS: The analysis revealed that IBL has a large and significant effect on enhancing students' critical thinking (g = 1.336; 95% CI [1.061, 1.611]). However, substantial heterogeneity was observed (I 2 = 92.09%), suggesting variability across contexts. Moderator analyses revealed that the main moderator, control group variations, was statistically significant in moderating the effectiveness of IBL (Qm = 5.21; p = .022). in contrast, subject domain (Qm = 1.43; p = .698), education level (Qm = 1.11; p = .774), and country ( Q m  = 3.33; p = .650), were insignificantly moderating the effectiveness of inquiry learning. CONCLUSIONS: The present meta-analysis highlighted that IBL is effective in improving students' critical thinking. However, the effectiveness of IBL was relative to the type of control group variations. Its effect on critical thinking was greater when compared with teacher-centered learning but smaller when compared with other student-centered learning.

Thinking

Revealing potential biomarkers and metabolic mechanisms of ovarian aging in hens during late laying period based on machine learning and metabolomics.

Ovarian function decline during the late laying period represents a major bottleneck for the economic efficiency of the global poultry industry. However, the underlying metabolic mechanisms and reliable early-warning biomarkers for ovarian aging remain poorly understood. In this study, we performed the first untargeted LC-MS/MS metabolomics analysis of ovarian tissues from Taihe silky fowls at peak laying (30 weeks) and late laying (50 weeks) stages, and employed an ensemble machine learning strategy integrating LASSO, random forest, and support vector machine (SVM) algorithms to identify high-confidence core biomarkers of ovarian aging. Gene expression analysis was further conducted to validate the potential molecular mechanisms. Our results showed that the metabolic profiles of ovarian tissues differed significantly between the two groups. A total of 6 core biomarkers were identified, 4 of which were long-chain acylcarnitines. Mechanistic analysis revealed that downregulation of key genes in the carnitine shuttle system led to impaired mitochondrial fatty acid β-oxidation, which in turn triggered excessive oxidative stress and compromised ovarian endocrine function. In conclusion, this study identifies long-chain acylcarnitines as potential metabolic biomarkers for ovarian aging in Taihe silky fowls. These findings provide novel insights into the metabolic basis of poultry ovarian aging and lay a theoretical foundation for the precise regulation of reproductive performance in indigenous poultry breeds.

Animals

Improving insurance deduction identification: a hybrid artificial intelligence model using machine learning and expert systems.

PURPOSE: Financial challenges in healthcare systems worldwide, especially in low- and middle-income countries like Iran, have increased hospitals' reliance on insurance reimbursements. Unrecognized insurance deductions often cause severe financial shortages, making efficient deduction management crucial. This study aimed to design a hybrid intelligent system for identifying and predicting insurance deductions by combining machine learning and expert system frameworks. DESIGN/METHODOLOGY/APPROACH: A mixed-methods design was applied in four stages. First, a scoping review identified the causes and patterns of insurance deductions. Second, interviews with 15 insurance experts produced a validated checklist and a dataset from inpatient billing records. Third, using the CRISP-DM methodology, machine learning algorithms were developed and tested in SPSS Modeler alongside a fuzzy expert system developed in MATLAB. Finally, the model was validated using the holdout method. FINDINGS: Four categories of deduction drivers were identified: service provision, registration errors, document submission issues, and revenue conversion processes. The CHAID decision tree outperformed other algorithms with a 99% precision rate and the lowest Mean Absolute Error (9.43). A brief assessment of potential overfitting was conducted to ensure that the CHAID model's high accuracy was interpreted cautiously and supported by the validation results. The fuzzy expert system with validated rules was adaptable for deduction classification, especially for cases unsuitable for quantitative modeling. ORIGINALITY/VALUE: The hybrid model improves detection and prevention of deductions, offering actionable insights for hospital administrators, insurers, and policymakers. Its implementation can enhance hospital information systems, streamline claims processing, and optimize revenue management amid financial constraints.

Machine Learning

Response-optimised training improves learning of a complex motor task and closely related motor tasks.

Regular physical exercise is essential for promoting healthy aging and longevity. In older adults with varying physical and cognitive decline, optimising exercise interventions is crucial to maximise benefits. A promising approach to achieve this goal is by adjusting task demands to individual abilities in turn preventing over- or underloading their abilities. In the field of motor learning, it is currently unclear whether such an optimised training improves not only performance on the trained task but also transfers to untrained motor and cognitive tasks. We conducted a randomized, single-blinded, 6-week dynamic balance training (DBT) with healthy older adults (n = 30). Training was tailored to individual balance ability. Participants were assigned to either suboptimal (high or low difficulty) or optimal (moderate difficulty) training groups. Transfer effects were assessed via cognitive tasks (memory and executive) and motor tasks (untrained DBT variations and other balance tasks) measured pre-, mid- and post-intervention. Multivariate longitudinal statistical analysis showed higher performance gains in the optimal training group in three out of six sessions compared to the suboptimal groups, especially under testing conditions with high task demands. The optimal group also showed greater improvements in near motor transfer tasks mid- and post-intervention, while no significant differences were observed in the cognitive tasks. Within-group DBT learning positively correlated with transfer gains, highlighting the role of training response in achieving transfer. In conclusion, optimised task difficulty in balance training enhances both task-specific performance and related motor skills, supporting the use of personalised interventions to maintain function and independence in older adults.

Humans

Predicting ACL injury risk in athletes: A systematic review of machine learning-based models.

BACKGROUND: Early ACL injury risk identification in athletes is essential. This systematic review examines machine learning (ML) models for predicting ACL injuries, evaluating their methodological quality, performance, and reliability. METHOD: A comprehensive electronic search was conducted across PubMed, Scopus, Web of Science, and IEEE Xplore databases, supplemented by Google Scholar for grey literature, covering articles published between January 1, 2015, and August 30, 2025. Eligible studies were appraised using the Prediction Model Study Risk of Bias Assessment Tool (PROBAST) for methodological quality and risk of bias, and the Transparent Reporting of a Multivariable Prediction Model for Individual Prognosis or Diagnosis (TRIPOD) guidelines for quality of evidence. RESULTS: Ten studies were included. PROBAST showed eight studies had moderate risk of bias and two low risk. TRIPOD found only two studies met quality criteria. ML models included logistic regression (n = 5), support vector machines (n = 4), k-nearest neighbor (n = 3), decision trees (n = 3), random forests (n = 5), neural networks (n = 2), linear discriminant analysis (n = 1), and pre-trained CNNs (n = 1). AUC ranged from 0.63 to 0.98. Accuracy (reported in six studies) ranged from 26% to 95%; however, these values should be interpreted with caution due to the absence of confidence intervals, lack of class imbalance handling, and limited external validation across studies. Tree-based ensemble methods such as random forest achieved competitive accuracy (74-86%), while SVM, a non-ensemble classifier, reported accuracy ranging from 71% to 95%; however, the highest values were obtained in studies with notably small sample sizes (n = 12 to n = 39), raising concerns about overfitting and generalizability. CONCLUSION: Current ML algorithms show promise for identifying athletes at high ACL injury risk and detecting relevant risk factors. Although study quality was generally satisfactory, future research should prioritize external validation and model interpretability to support clinical translation.

Humans

Not just when, but how: An exploratory dual-control approach to video feedback in motor learning.

The present study provides exploratory evidence for a novel dual-control paradigm. It examines whether combining temporal over video feedback timing with learner-controlled interactive playback functions (pause, slow-motion, rewind) would enhance motor skill acquisition beyond temporal autonomy alone. Sixty-four novice adults were randomly assigned to one of four conditions: Full Control (self-controlled timing + interactive replay), Partial Control (self-controlled timing + non-interactive replay), Yoked Full Control (externally controlled timing + interactive replay), or Yoked Partial Control (externally controlled timing + non-interactive replay). Motor accuracy (Radial Error), movement consistency (Bivariate Variable Error), technical execution, and self-efficacy were assessed at pre-test, 24-h retention, and 72-h retention following two acquisition sessions on a dart-throwing task (120 trials total). The Full Control group demonstrated the greatest and most durable learning gains across all outcomes. The Group × Time interaction was significant across all dependent variables (η2ₚ ranging from 0.140 to 0.234), with Full Control demonstrating superior retention at both 24 and 72 h relative to other groups (though differences relative to Partial Control were more pronounced at 72-h retention). Critically, the Yoked Full Control group showed comparatively weaker outcomes despite access to the same interactive playback functions. These findings suggest that interactive video tools may be most useful when learners can regulate both when feedback is accessed and how it is inspected. Theoretical and practical implications for the design of learner-centered video feedback systems are discussed.

Humans

Machine learning-ready genomic biomarkers: ATF3 polymorphisms predict postoperative analgesic demand through AI-compatible phenotyping.

PURPOSE: To determine whether ATF3 polymorphisms can serve as genetic biomarkers for machine learning-based precision analgesia by establishing a genotype-phenotype association suitable for predictive modeling of postoperative opioid requirements. METHODS: In a prospective cohort of 167 adults undergoing abdominal surgery, ATF3 SNPs rs3122721 and rs3125293 were genotyped. A structured dataset architecture was developed to represent genetic profiles as input features for supervised learning models, enabling translational analysis of genotype‑dependent opioid consumption over 72 h. RESULTS: Patients with homozygous genotypes of the ATF3 SNPs had significantly higher opioid requirements than non‑carriers, despite reporting similar subjective pain scores. This consistent genotype‑dependent pattern provided a clinically relevant phenotype suitable for integration into predictive algorithms. CONCLUSION: ATF3 genotyping offers a promising biomarker for computationally informed precision analgesia. By linking genomic variability to clinically meaningful outcomes within a structured clinical and genomic framework, this approach supports the future development of risk-stratified clinical decision-support systems to optimize postoperative pain management.Trial registration ChiCTR1900021991, registered 30 April 2019. SUPPLEMENTARY INFORMATION: The online version contains supplementary material available at https://doi.org/10.1007/s13755-026-00480-9.

ATF3

Predicting training outcomes for developmental dyslexia from EEG data.

Developmental dyslexia (DD) is characterised by lower-than-average reading abilities and is diagnosed in approximately 10% of individuals. The societal barriers may limit professional fulfilment and psychological wellbeing of individuals with DD, calling for the development of effective interventions to counteract them. As DD is associated with challenges in both phonological and visuo-attentional domains, different longitudinal training approaches were developed to strengthen them. However, they require a considerable amount of personal, social and economic resources and the outcomes may vary depending on individual differences in behavioural and neurophysiological functionality. Hence, predicting training outcomes might help in developing personalised treatment protocols and optimising the use of resources. In the present work we applied machine learning to resting-state EEG to predict longitudinal training outcomes in adults with DD enrolled in a randomized clinical trial. In particular, one group received a visuo-attentional training combined with transcranial alternating current stimulation (tACS), another group received visuo-attentional training with sham/placebo stimulation, and the third group received a phonological training with sham/placebo stimulation. The improvement in text reading speed was associated with spectral power in low-beta and individual frequencies in the alpha (IAF) and beta (IBF) bands, while the improvement in pseudoword reading was associated with IBF. The findings highlight the potential of capturing neural markers of treatment responsiveness in DD. Future studies should focus on the generalisability of predictive models to real-world settings, while investigating whether specific EEG markers predict responsiveness to distinct remediation protocols, thus supporting the development of personalised interventions.

Humans

The role of simulator immersion on learning and transfer of decision-making skill in sport.

Virtual reality has become popular in sport and other domains because it can immerse the user within a sporting context and solve logistical problems for additional off-field training. There is limited evidence, however, of whether immersion is crucial for learning and transfer. This study compared training of decision-making skill between 360-degree video virtual reality (360VR) and two-dimensional video. Twenty-eight Australian Rules Football players were randomly assigned to one of three training groups: 360VR, two-dimensional video, and control. Across four weeks, participants in the training groups were exposed to decision-making scenarios consisting of visual, contextual and auditory cues. Performance was assessed pre- and post-training with virtual reality and field-based decision-making tests. Results indicated that the two-dimensional video training group showed significantly superior decision-making in the field-based transfer test compared to 360VR and control groups post intervention. There was also indication that two-dimensional video training was superior to the control post intervention in the virtual reality test. Findings indicate that immersion created in virtual reality is not an underpinning mechanism for learning and transfer, rather the use of perceptual information is crucial. 360VR may facilitate uptake through engagement, but two-dimensional video is adequate for learning and transfer of decision-making to the field.

Humans

Mul-PheG2P: decoupled learning and prediction-space fusion enables robust and interpretable multi-phenotype genomic prediction.

Genomic prediction of multiple phenotypes is crucial in modern plant breeding; however, existing methods struggle with negative transfer and lack interpretability, particularly across high-dimensional small-sample data and diverse species. To address this, we propose Mul-PheG2P, a novel paradigm based on decoupled learning and predictive space fusion. It employs a two-stage design: first training phenotype-specific encoders using genetic data, then decoupling phenotype-specific learning from cross-phenotype aggregation via an interpretable prediction layer. Mul-PheG2P outperforms existing methods across diverse crop datasets, including maize (Zea mays), wheat (Triticum aestivum), and tomato (Solanum lycopersicum). It provides a multi-scale interpretability chain: at the macro level, it quantifies phenotypic contributions via attention-based weighting; at the micro level, Integrated Gradients reveal the genetic basis of predictions. Notably, the model successfully identified the CCT (CONSTANS, CO-like, and TOC) motif regulating photoperiodism and the SQUAMOSA (SQUAMOSA promoter binding protein) promoter for inflorescence development, confirming its ability to capture functional biological mechanisms. These results highlight the high performance and interpretability of Mul-PheG2P, showcasing its value for low-cost, large-scale screening to advance precision breeding.

Phenotype

Future promise, current clinical ambiguity: a systematic review of machine learning algorithm outputs predicting risk of cardiovascular disease.

OBJECTIVE: To examine whether the outputs of machine learning algorithms designed to predict risk of cardiovascular disease (CVD) address known deficiencies of the Framingham Risk Score (FRS) and improve risk estimates. METHODS: For this critical review, Medline, Embase and IEEE were searched from inception to 1 January 2025. Included were studies describing machine learning algorithms designed to specifically compare output of cardiovascular risk assessment with the FRS. Commentaries, letters, unpublished work or non-peer-reviewed papers were excluded.Following Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) guidelines, two reviewers screened titles and abstracts independently, then populated a purpose-built data extraction form. A subsequent qualitative thematic analysis focused on algorithms' strengths, added value, potential harms, unintended consequences and equity implications.The main outcome assessed was whether, among healthy adults, the algorithm improved CVD risk prediction relative to the FRS. RESULTS: Of 707 studies retrieved, 29 met inclusion criteria. 23 reported improved predictive ability relative to the FRS. Most datasets and/or medical records used included sociodemographic predictors of CVD not included among FRS inputs. Some added costly diagnostic tests like CT angiography to FRS screening indicators. When they were defined, inputs and outcomes such as hypertension or myocardial infarction did not always adhere to FRS values. Statistical significance was generally taken as a proxy for clinical significance. Some algorithms overestimated the number at risk compared with the FRS without discussing whether that larger proportion might be at risk of overdiagnosis rather than CVD, while a few decreased the proportion found to be at risk. CONCLUSIONS: Use of artificial intelligence to improve accuracy of risk assessment for CVD demonstrates the technological capacity to merge known sociodemographic predictors with biologic variables and examine non-linear interactions among these. Still needed to achieve patient benefit is clinical insight, adherence to screening principles and cost-benefit assessment of inputs selected.

Humans

Assessment of the Potential of Different Anthropometric Indices in Predicting the Risk of Diabetes and Associated Co-morbidities.

Diabetes, a chronic disorder, is showing a rapidly increasing trend globally. India holds the second position in the global diabetes epidemic. The present investigation is an assessment of different anthropometric measurements and their association with type 2 diabetes to determine their diagnostic potential for diabetes as well as its co-morbidities. In this cross-sectional study, we have measured anthropometric parameters and blood biomarkers in subjects with diabetes. We have presented the comparisons of cost- and time-effective anthropometric variable with costly and time-dependent biochemical variables in control and diabetic groups (n = 233/group). Correlations between anthropometric variables and biochemical measurements, as well as the diagnostic utility of anthropometric variables for diabetes, were evaluated. The diagnostic utility of anthropometric variables for diabetes was assessed through receiver operating characteristic (ROC) curves. Neck circumference, sagittal abdominal diameter (SAD), skinfold thickness, and body roundness index (BRI) displayed high specificity and diagnostic utility for diabetes, emphasizing their potential in predicting diabetes and the further development of metabolic syndrome. The study highlights the importance of cost- and time-effective anthropometric assessments in diabetes risk evaluation and calls for further research to elucidate this intricate relationship and develop personalized management strategies.

Humans

RR-interval-based atrial fibrillation detection and burden estimation: cross-dataset validation and calibration-aware probability analysis.

Objective.Atrial fibrillation (AF) burden has become an increasingly important endpoint in long-duration rhythm monitoring, but reliable burden estimation requires more than accurate AF detection alone. In particular, when burden is derived by aggregating predicted AF probabilities over time, probability calibration may directly affect burden validity under external dataset shift.Approach.This study developed an interpretable-interval feature model for AF detection and evaluated it using record-wise cross-validation on a development cohort and independent cross-dataset external validation on public Holter electrocardiographic databases. Window-level performance was assessed using the area under the receiver operating characteristic curve (ROC-AUC), area under the precision-recall curve (PR-AUC), Brier score, expected calibration error (ECE), and calibration intercept and calibration slope. Recording-level AF burden was estimated using both probability-based and hard-label aggregation and evaluated using mean absolute error (MAE) and agreement analyses.Main results.The model showed high discrimination in both development and external evaluation, with external ROC-AUC ofand PR-AUC of. However, external calibration deteriorated despite preserved ranking performance, with Brier score of, ECE(15) of, calibration intercept of, and calibration slope of. In the external cohort, probability-based burden estimation preserved strong association with reference burden but showed weaker raw agreement than hard-label aggregation, with MAE ofversus, consistent with systematic probability underprediction. Repeated external recalibration across record-level splits substantially improved probability quality and probability-based burden estimation. Median probability-burden MAE decreased fromwithout recalibration toafter Platt recalibration andafter isotonic recalibration, while median ECE(15) decreased fromtoand, respectively.Significance.These findings indicate that-interval-based AF detection maintained strong ranking performance in the tested external cohort, but probability calibration should be evaluated explicitly when predicted probabilities are aggregated into AF-burden estimates.

Atrial Fibrillation

Meta-PseU: A meta-classifier for robust prediction of RNA pseudouridine modification sites from long sequences.

BACKGROUND AND OBJECTIVES: Pseudouridine (Ψ) represents one of the most abundant and conserved RNA modifications. Ψ provides an additional hydrogen-bond donor that enhances RNA structural stability and modulates translation. It participates in diverse biological processes, including RNA-protein interactions, splicing, translational control, and stress responses. Aberrant pseudouridylation is implicated in cancer, neurodegenerative disorders, and autoimmune diseases. Despite its biological importance, experimental identification of Ψ sites remains time-consuming and costly, limiting the feasibility of transcriptome-wide profiling. Computational approaches have therefore become essential complements to experimental techniques. However, state-of-the-art machine-learning and deep-learning predictors often suffer from limited generalizability due to small training datasets. To overcome these issues, we aim at constructing new long-sequence datasets and developing a novel Ψ site predictor. METHODS: New long-sequence datasets were constructed as benchmarks for RNA Ψ-site prediction. The Ψ modification sites in RMBase 3.0 were mapped to the reference genomes across three species of human, mouse, and yeast, and the RNA sequences with a length of 201 were generated by extending the upstream and downstream from the mapped, central sites. To eliminate sequence redundancy, the sequences were clustered using CD-HIT with a 70% sequence identity threshold. We developed Meta-PseU, a logistic regression-based meta-classifier that considered 118 machine learning and deep learning classifiers. The datasets and programs are freely accessible at https://github.com/kuratahiroyuki/MetaPseU. RESULTS: By optimizing model configuration, we proposed the Meta-PseU model stacking 32 machine learning and deep learning classifiers out of 118 classifiers. Meta-PseU substantially improved model generalizability, overcoming a key limitation of existing approaches. It greatly outperformed state-of-the-art predictors and achieved increasing accuracy with increasing sequence length. CONCLUSIONS: Long-sequence datasets were newly constructed as benchmarks for RNA Ψ-site prediction. Meta-PseU offers a new framework for robust Ψ-site identification by using long sequences.

Pseudouridine

Data-centric, robust, and explainable multimodal deep learning for clinical decision support: A systematic review.

PURPOSE: Multimodal deep learning is increasingly proposed for clinical decision support (CDS) under a "data-centric" framing that prioritizes label quality, missing-modality robustness, distribution shift, calibration, and explainability. Prior reviews have examined multimodal medical AI, CDS, and data-centric methods separately, but none address their intersection. We mapped the modalities, fusion strategies, and data-centric and explainability techniques used in this recent literature, quantified how often each is implemented rather than merely mentioned, assessed deployment-relevant evidence (external validation, clinical-outcome measurement, equity), and formally appraised study-level risk of bias. METHODS: Following the PRISMA 2020 statement (PROSPERO CRD420261427815; registered retrospectively), we screened 150 records and included primary, clinical, multimodal studies that applied machine or deep learning to a decision-support task and reported at least one quantitative result. Two reviewers screened and extracted data with consensus adjudication. Each study was coded against pre-specified operational definitions, separating implemented or empirically evaluated techniques from those only mentioned. Study-level risk of bias was assessed with PROBAST + AI. Synthesis was narrative. RESULTS: Thirty-one studies met inclusion; 30 (97%) were published between 2024 and 2026, with a median of three modalities (range 2-6), most commonly structured EHR (71%) and imaging (39%). Data-centric techniques were frequently reported (74-84% across label-noise, distribution-shift, calibration, missing-modality and class-imbalance handling; equity 61%). However, external validation was reported in only 4/31 studies (13%), a clinical or provider outcome in 3/31 (10%), and no study reported routine deployment. Overall risk of bias was high in 27/31 studies (87%), driven by the analysis domain. CONCLUSION: Within this recent, self-selected slice of the field, technical robustness and explainability techniques are widely reported but rarely validated out-of-distribution or against clinical outcomes, and the underlying evidence is at high risk of bias. Progress requires external multi-site validation, clinical-outcome measurement, formal bias appraisal, and adherence to AI reporting standards (e.g., TRIPOD + AI) before deployment can be justified.

Deep Learning

A framework for delivering real-time, instrument-relative navigation in transoral robotic surgery.

Transoral robotic surgery (TORS) is a minimally invasive, inside-out technique that, compared with traditional open approaches, provides fewer post-operative complications, shorter hospital stays, and improved survival for early-stage head and neck cancer. However, TORS is limited by its steep learning curve and poor visualization of deep tumor margins. This randomized crossover study evaluated a surgical navigation system's potential to enhance accuracy and user experience with real-time, instrument-relative feedback. Seven Teflon beads (d = 2.381 mm) were embedded at the tongue base of a porcine pharynx-and-larynx model. Tongue blade compression and retraction were applied to the model to mimic intraoperative tissue deformation, reproducing the anatomical shifts that occur relative to preoperative imaging. Eight participants used the da Vinci Surgical system to localize the beads by placing pins under two conditions: (a) preoperative computed tomography with no navigation; (b) model-based visual navigation with quantitative instrument-to-target metrics. Surgical accuracy was determined by calculating the target localization error (TLE, pin-to-bead Euclidean distance) and the angular error (AE, pin axis trajectory to bead). Accounting for training level and bead depth, surgical navigation reduced TLE by 5.44 mm (95% CI, 4.02-6.86 mm; p = 2.00e-11) and AE by 8.47 degrees (95% CI, 6.21-10.72 degrees; p = 5.17e-11). Impressions of the system were generally favorable using a 5-point Likert survey and task duration (p = 0.26) or cognitive workload via the NASA-Task Load Index (p = 0.22) were not significantly affected. The navigation system demonstrated translational promise, offering improved target localization accuracy and more consistent performance across experience levels, two critical determinants of surgical quality in TORS.

Robotic Surgical Procedures