Search PubMedSearch

PubMed · 40989216

Accurate identification of abnormal ploidy using an artificial intelligence model in preimplantation genetic testing.

Abstract

STUDY QUESTION: Can ultra-low-coverage whole-genome sequencing (ulc-WGS) accurately identify abnormal ploidy during preimplantation genetic testing (PGT)? SUMMARY ANSWER: The artificial intelligence (AI)-based PGT-Plus model demonstrates high accuracy in ploidy detection, offering a cost-effective solution that enhances clinical utility of PGT. WHAT IS KNOWN ALREADY: The predominant PGT for aneuploidy can identify chromosomal aneuploidies but cannot determine ploidy status. Transferring embryos with ploidy abnormalities can result in miscarriage and molar pregnancy. On the other hand, in ART, fertilization is assessed by morphological pronuclear assessment at the zygote stage. However, it has a low specificity in the prediction of abnormal ploidy status and embryos deemed abnormally fertilized can yield healthy pregnancies. Accurately identified abnormal ploidy in PGT-A can resolve current limitations and expand the utility range of PGT-A. Several studies have identified ploidy abnormalities; however, they were mainly based on single-nucleotide polymorphism (SNP) arrays or needed to combine additional targeted-next-generation sequencing (NGS) information. Studies based on ulc-WGS remain scarce. STUDY DESIGN SIZE DURATION: The study consisted of two stages: methodology establishment and validation. An AI model, named PGT-Plus, was developed using 653 samples with known ploidy status, which was further validated using 792 different ploidy status samples. In the clinical application stage, the approach was used to analyse the ploidy status of 19&#x2009;103 normally fertilized PGT blastocysts and 140 single pronucleus (1PN)-derived blastocysts collected between May 2022 and December 2023. All blastocysts were tested using trophectoderm biopsy and NGS. PARTICIPANTS/MATERIALS SETTING METHODS: The methodology is based on the ulc-WGS data. First, based on samples with known ploidy status: the heterozygosity rate of high-frequency biallelic SNPs, the likelihood ratio (LLR) of alleles was calculated under different assumptions ('both parental homologs' [BPH] from a single parent, 'single parental homolog' [SPH] from each parent, disomy, and monosomy) by leveraging allele frequencies and linkage disequilibrium (LD) measured in the 1000 genomes project database. Twenty-three continuous candidate features derived from heterozygosity rates and LLRs of chromosomes or selected windows were included to establish the ploidy prediction AI model. Gini importance analysis and multicollinearity mitigation was performed for feature selection, then the performance of Random Forest (RF), Support Vector Machine (SVM), and Logistic Regression for modelling was compared. Subsequently, the parameter optimization was performed based on the RF model. Ploidy constitution concordance was evaluated in known ploidy status samples. The frequency of abnormal ploidy in normal fertilized PGT blastocysts and 1PN-derived blastocysts (including conventional IVF and ICSI) was evaluated. MAIN RESULTS AND THE ROLE OF CHANCE: Eleven features were collected for model architecture compared to SVM and Logistic Regression; RF achieved superior performance for ploidy detection. The AI model achieved an AUC of 1 for genome-wide-uniparental diploidy (GW-UPD), 1 for triploidy, and 0.99 for diploidy. For the 792 validation samples, 99.5% of samples were successfully detected using the AI model, and the model showed 100% accuracy for ploidy classification. In the clinical application stage, out of 19&#x2009;103 PGT samples, 19&#x2009;069 were successfully analysed using the model, with 110 (0.57%) identified as having abnormal ploidy embryos. Among these, 12.7% (14/110) were identified as GW-UPD, and 87.3% (96/110) were triploid. Among 5563 diploid blastocysts transferred, 3478 clinical pregnancies were achieved. Subsequent ploidy analysis was performed for 217 spontaneous abortion and 935 prenatal diagnostic samples, and no abnormal ploidy was identified. Furthermore, of the 140 1PN embryos tested, 40 (28.6%) exhibited GW-UPD, 3 (2.1%) exhibited triploidy, and 97 (69.3%) were determined to be biparental and normally fertilized. Among the 97 biparental embryos, 46 were diploid, 11 were mosaic, and 40 were aneuploid. In terms of the insemination pattern, the percentage of abnormal ploidy in ICSI was significantly higher than in conventional IVF (P&#x2009;<&#x2009;0.01, 37.1% vs. 2.9%, respectively). With full informed consent, 20 patients without euploidy from normal fertilization chose 1PN-derived biparental and diploid blastocysts to transfer, resulting in 10 clinical pregnancies and 9 ongoing pregnancies. LARGE-SCALE DATA: N/A. LIMITATIONS REASONS FOR CAUTION: Some rare ploidy abnormalities, such as polyploidy with an equal number of identical sets of chromosomes and ploidy mosaicism cannot be accurately identified. Moreover, the origin of abnormal ploidy was not identified due to the unavailability of DNA from both parents. WIDER IMPLICATIONS OF THE FINDINGS: The PGT-Plus AI model provides a ploidy evaluation method based on the conventional PGT-A data and integrates directly into standard PGT-A workflows. Clinical utility results suggest that the model is a valuable tool for identifying embryos with abnormal ploidy in PGT-A and rescuing normal diploid embryos from abnormally fertilized embryos. These findings demonstrate that PGT-Plus significantly enhances the diagnostic accuracy of PGT. STUDY FUNDING/COMPETING INTERESTS: This study was supported by grants from Major Scientific Program of CITIC Group (No. 2023ZXKYB34100, to Ge.L.), Hunan Provincial Grant for Innovative Province Construction (2019SK4012), Hunan Xiangjiang New District (Changsha High-tech Zone) key core technology research project in 2023, and Science Foundation of Hunan Province (Grant 2023JJ30422). All authors declared no conflicts of interest..

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Pingyuan Xie, Rijing Pang, Luyao Zeng, Shuoping Zhang, Lei Sun, Kaisen Yang, Xiaoyi Yang, Shuang Zhou, Senlin Zhang, Guangjian Liu, Yueqiu Tan, Liang Hu, Fei Gong, Jia Fei, Ge Lin. 2025-09-02. Accurate identification of abnormal ploidy using an artificial intelligence model in preimplantation genetic testing.. https://doi.org/10.1093/hropen%2Fhoaf054

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The next-generation biomarkers in early-stage triple negative breast cancer.

PURPOSE OF REVIEW: Despite maximal neoadjuvant chemoimmunotherapy, nearly 40% of early-stage triple-negative breast cancer (TNBC) patients fail to achieve pathological complete response, underscoring an urgent need for biomarkers capable of guiding treatment modulation. This review summarizes recent advances in tumor-infiltrating lymphocytes (TILs), circulating tumor DNA (ctDNA), and genomic signatures, exploring their potential integration into clinical decision-making. RECENT FINDINGS: TILs remain the most validated, cost-effective prognostic biomarker, although standardized scoring is still needed to overcome interobserver variability. ctDNA has emerged as a dynamic, real-time prognostic tool, with postneoadjuvant detection strongly predicting relapse and worse outcomes. Nonetheless, optimal sampling timing remains undefined. Genomic signatures, particularly TNBC-DX, provide standardized, reproducible prognostic information by integrating immune and proliferative gene expression. Emerging data suggest that combining these biomarkers may offer a complementary and synergistic effect. SUMMARY: Multibiomarker integration, supported by prospective validation and automated models, represents a promising approach to personalize treatment algorithms in early-stage TNBC, balancing efficacy and toxicity while guiding escalation and de-escalation strategies.

artificial intelligence

Mechanistic Perspectives From Genomics and Pangenomics of Medicinal and Aromatic Plants: Linking Genome Architecture to Phytochemical Diversity.

Medicinal and aromatic plants (MAPs) produce a remarkable diversity of specialized metabolites with significant pharmaceutical, nutraceutical, and industrial value. Although advances in long-read sequencing, chromosome-scale genome assembly, and pangenomics have greatly expanded genomic resources, the mechanistic links between genome architecture and phytochemical diversity remain incompletely understood. The present review synthesizes current evidence describing how structural genomic variation may contribute to phytochemical diversity, while acknowledging that many proposed genome-to-metabolite relationships require further experimental validation. Examples illustrate how genome architecture is associated with specialized-metabolite biosynthesis through multiple regulatory processes. However, the strength of supporting evidence varies considerably among MAP species. Moreover, relatively few genome-to-metabolite relationships have been confirmed through direct functional validation. We further discuss how pangenomics, multiomics integration, genome editing, synthetic biology, and artificial intelligence support the discovery, validation, and engineering of specialized metabolic pathways. Casual conclusions are evaluated according to the strength of available evidence, highlighting where causal relationships have been experimentally established and where conclusions remain primarily association-based. Overall, this review provides an integrated conceptual and evidence-based perspective summarizing proposed relationships between genome architecture and phytochemical diversity and outlines future priorities for functional genomics, precision breeding, metabolic engineering, and sustainable utilization of MAPs.

artificial intelligence

Digital and computational morphology in hematology: current platforms, clinical evidence, and future requirements.

INTRODUCTION: Morphologic examination of peripheral blood and bone marrow remains central to the diagnosis and classification of hematologic disorders. Conventional optical microscopy, however, is labor-intensive, dependent on operator expertise, and affected by interobserver variability. Digital morphology has developed from automated image acquisition and cell pre-classification into a broader field that includes whole-slide imaging, remote review, quantitative morphometry, and artificial intelligence-based analysis. CONTENT: This review examines current applications of digital morphology in peripheral blood, bone marrow aspirates, malaria detection, and body-fluid analysis. Commercial platforms are evaluated with particular attention to the distinction between raw automated pre-classification, expert digital post-classification, and comparison with independent optical microscopy. Digital systems generally perform well for common mature leukocyte populations but remain less reliable for rare or diagnostically critical cells, including blasts, abnormal lymphoid cells, plasma cells, and intermediate maturation stages. Research systems increasingly extend analysis from individual-cell classification to whole-slide, specimen-level, and patient-level assessment. SUMMARY: Digital morphology can improve standardization, image traceability, remote consultation, education, proficiency testing, quality assurance, and selected aspects of laboratory workflow. Its clinical value depends on appropriate validation, transparent reporting of reference methods, recognition of algorithm-specific failure modes, and clearly defined criteria for expert review and conventional microscopy. Human expertise remains essential not only for validating results but also for adapting cell taxonomies and interpretive rules to evolving classifications of hematologic diseases. OUTLOOK: Future progress will require representative multicenter datasets, harmonized morphologic terminology, external validation, interoperability with laboratory information systems, and continuous monitoring after software or hardware updates. Integration of morphology with quantitative hematology, flow cytometry, cytogenetics, genomics, and clinical data may support more comprehensive computational diagnosis. Digital platforms may also broaden access to specialist expertise, training, and quality programs in resource-limited institutions and regions, provided that infrastructure, governance, and professional competency are adequately supported.

artificial intelligence