Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Multi-omics data integration”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6Linked to original sources

Bridging the gap: multi-omic insights into exercise responses in postmenopausal women.

Postmenopausal women represent the fastest-growing demographic at risk of sarcopenia and cardiometabolic disease, yet exercise biology research remains disproportionately derived from male or hormone-replete phenotypes. Menopause constitutes a chronic endocrine perturbation characterized by sustained reductions in estrogen and progesterone, and altered androgen balance, superimposed on the acute and chronic perturbations induced by exercise. This hormonal shift modifies substrate metabolism, inflammation, redox balance, and recovery capacity, factors that shape molecular responses to exercise across tissues and time. Here, we synthesize current evidence on exercise responses in postmenopausal females across genomics, epigenomics, transcriptomics, proteomics, and metabolomics/lipidomics. Across omics layers, direct data in postmenopausal cohorts remain limited, with frequent underreporting of menopausal status, hormone therapy exposure, circulating hormone concentrations, medication use, and biosampling timing relative to exercise and hormone dosing. We outline a menopause-aware framework for exercise-omics that prioritizes endocrine stratification, repeated sampling across exercise and recovery, and integrative multi-omics approaches linking molecular responses to functional outcomes. We also outline minimum reporting standards to improve reproducibility, inclusivity, and translational relevance. Advancing menopause-aware exercise-omics will be essential for developing precision exercise strategies that improve health span and functional independence in later life.

Humans↗

Multimodal artificial intelligence and machine learning in oncology: from data integration to precision cancer care.

Cancer remains a major global health burden, with approximately 20 million new cases and 9.7 million cancer-related deaths reported globally in 2022. While advances in radiological imaging, molecular profiling, and clinical data have enhanced the interpretation of disease progression, the availability of multiple such modalities still does not meet the needs of a large patient population. This narrative review focuses on the role of multimodal artificial intelligence and machine learning in bridging the gap in interpreting heterogeneous modalities to improve risk prediction, prognostic assessment, and treatment decision-making in precision oncology. Multimodal frameworks such as Pathomic Fusion illustrate how complementary histopathological and genomic information can be integrated for cancer diagnosis and prognostic modeling. Multimodal models have demonstrated potential in virtual biopsy, cancer screening, prognostic prediction, radiotherapy planning, intraoperative guidance, and clinical-trial design using digital twins and synthetic control arms. The major limitations of incorporating multimodal artificial intelligence and machine learning in oncology include data heterogeneity, demographic or institutional biases, and reproducibility challenges that hinder translation. Accordingly, appropriate data-governance strategies, fairness audits, and privacy-preserving approaches such as federated learning should be considered where appropriate. Future progress will depend on the development of standardized benchmarking datasets, robust external validation, seamless integration with electronic health records and picture archiving and communication systems, and the implementation of explainable, secure, and clinically validated multimodal artificial intelligence frameworks that support precision oncology in routine clinical practice.

deep learning↗

Multi-omics Approaches to CCAAT/Enhancer-Binding Protein Beta in Oral Squamous Cell Carcinoma: Crosstalk Between Tumor Cells and Tumor-Associated Macrophages Driving Disease Progression.

BACKGROUND: CCAAT/Enhancer-Binding Protein Beta (CEBPB) is an important transcription factor that regulates tumor progression. However, the mechanism by which CEBPB regulates the progression of Oral Squamous Cell Carcinoma (OSCC) remains incompletely understood. Tumor progression depends on complex intercellular interactions within the tumor microenvironment. The purpose of this study was to investigate the role and epigenetic regulatory mechanisms of CEBPB in interactions between OSCC cells and tumor-infiltrating immune cells. METHODS: Bulk RNA-seq, ChIP-seq, and scRNA-seq data were obtained from The Cancer Genome Atlas (TCGA) database and the Gene Expression Omnibus (GEO) database. The HOMER algorithm was employed to identify enhancers and predict the CEBPB-binding motif. Cell cluster analysis, functional enrichment, and intercellular interaction analysis were performed using the "Seurat" R package. H3K27ac enrichment at GAS6 enhancers was validated by ChIP-qPCR. Metastatic OSCC cells with CEBPB knockdown or GAS6 overexpression were established and co-cultured with THP-1 cells. IL-10 and IL-6 secretion from co-cultured THP-1 cells was detected via ELISA. Chemotaxis of OSCC cells toward THP-1 cells was assessed through a Transwell assay. RESULTS: CEBPB was upregulated in OSCC and correlated with poor prognosis. By integrating H3K27ac ChIP-seq and bulk RNA-seq data, 131 CEBPB-regulated enhancer-controlled genes were identified in lymph node metastatic OSCC cells. scRNA-seq analysis revealed eight major cell clusters in primary foci and lymph node metastases, including T/NK cells, malignant epithelial cells, B/plasma cells, macrophages, fibroblasts, dendritic cells, endothelial cells, and mast cells, with the malignant epithelial cells stratified into distinct sub-clusters. CEBPB expression was elevated in malignant epithelial cells of lymph node metastases compared to primary foci. Furthermore, 15 pairs of enhanced ligand-receptor interactions were identified in lymph node metastases relative to primary foci. GAS6 was a CEBPB-regulated enhancer-controlled gene, primarily mediating interactions between malignant cells and macrophages. CEBPB knockdown in metastatic OSCC cells significantly impaired their chemotaxis toward cocultured THP-1 cells, and downregulated IL-10/IL-6 secretion and CD206 expression in cocultured THP-1 cells. Conversely, GAS6 overexpression reversed these inhibitory effects. CONCLUSION: CEBPB activated GAS6 transcription in metastatic OSCC cells. The CEBPB/ GAS6 axis in metastatic OSCC cells enhanced their chemotaxis toward macrophages and promoted the M2 polarization of macrophages, thereby facilitating the establishment of an immunosuppressive microenvironment.

Humans↗

NMR metabolomics and glycomics for cancer detection in patients with non-specific symptoms: a prospective observational cohort study.

BACKGROUND: Early cancer diagnosis in patients with non-specific symptoms is limited by the lack of discriminatory tests. Within the Oxfordshire Suspected CANcer (SCAN) pathway, exploratory biomarker work showed that serum 1H NMR-based metabolomics can identify cancer with high accuracy. SCAN2 evaluated whether integrating metabolomics with glycomics provides complementary molecular information and improves discrimination in a clinically complex, real-world population. METHODS: Serum from 369 SCAN patients (59 cancers) was analysed using AXINON® System-derived NMR metabolomics and HPLC-MS glycomics. Machine-learning models were trained to predict cancer status, with performance assessed by receiver operating characteristic (ROC) analysis of pooled cross-validated predictions. To place cancer risk in a broader clinical context, a second classifier modelling alternative non-cancer diagnosis was incorporated, and mean predicted probabilities from both models were jointly projected into a two-dimensional space, maintaining strict separation of training and test data. FINDINGS: In the full cohort, integration of glycomics with metabolomics achieved an AUC of 0.814 (95% CI 0.808-0.820). In a refined sub-cohort excluding major comorbidities and selected cancer types (32 cancers, 277 non-cancers), performance improved to an AUC of 0.884 (95% CI 0.879-0.890). Discriminatory features included cancer-associated biantennary fucosylated glycans alongside amino acid metabolites (glutamate, histidine) and lipoprotein-related measures. A classifier distinguishing metastatic from non-metastatic disease (n = 29 vs. 30) achieved an AUC of 0.80. Joint probability analysis in the full cohort preserved cancer-associated signatures across comorbidity burden, with projection-based classification achieving an accuracy of 89.2% (95% CI 85.7-92.6). INTERPRETATION: These findings validate the SCAN1 metabolomic signature in a more clinically complex cohort and indicate that integrating glycomics with metabolomics provides complementary biological information for cancer discrimination. Joint probability analysis provides an interpretable framework for cancer risk stratification within multimorbid diagnostic pathways, supporting the clinical potential of scalable multi-omics blood testing. FUNDING: EPSRC, EU Horizon 2020, Wellcome/MLSTF, Novo Nordisk Foundation.

Humans↗

A Multi-omics Regulated Cell Death Framework Defines Immune Phenotypes and Guides Precision Therapy in Colorectal Cancer.

Colorectal cancer (CRC) is molecularly and immunologically heterogeneous, contributing to variable treatment response. Because regulated cell death (RCD) intersects with tumor metabolism, immune regulation, and therapeutic susceptibility, we built an RCD-centered framework for CRC stratification. Multi-cohort transcriptomic data were used to infer RCD subtypes with non-negative matrix factorization (NMF) and non-negative least squares (NNLS). Genomic, bulk RNA-seq, single-cell RNA-seq, and spatial transcriptomic datasets were integrated to characterize subtype-associated biology. Machine-learning models were developed for immunotherapy response and survival-risk estimation. Candidate compounds were screened by GDSC2-based drug-sensitivity modeling and molecular docking, and FSTL3 was functionally assessed in vitro. The framework separated CRC samples into two RCD-related phenotypes resembling immune-hot and immune-cold states. RCD1 showed immune activation and higher mutational burden, whereas RCD2 showed immune-suppressed features, intratumoral heterogeneity, and aggressive biology. RCD-associated signatures showed potential for predicting immunotherapy response and survival risk. Dasatinib was prioritized for immune-cold, high-risk tumors, with preliminary evidence supporting its activity in CRC cells, while functional assays suggested a role for FSTL3 in growth, invasion, epithelial-mesenchymal transition, and apoptosis regulation. These findings suggest that RCD-based multi-omics analysis may refine CRC stratification and help generate therapeutic hypotheses.

Colorectal cancer↗

Decoding age-stratified clinical and molecular heterogeneity in male breast cancer through multiomic profiling.

OBJECTIVE: Age-associated molecular heterogeneity is well described in female breast cancer but remains insufficiently characterized in male breast cancer (MBC). We profiled age-stratified clinical and molecular differences between younger (&#x2264;55 years) male breast cancer (YMBC) and older (>55 years) male breast cancer (OMBC). METHODS: We retrospectively analyzed 347 patients with MBC diagnosed at Fudan University Shanghai Cancer Center by integrating clinicopathological data, RNA sequencing, and whole-exome sequencing (WES). Survival, differential expression, and mutational signature analyses were performed. Tumor microenvironment features were inferred using xCell and ESTIMATE, and weighted gene co-expression network analysis (WGCNA) was conducted to identify age-associated co-expression modules. Candidate therapeutics were prioritized using the Genomics of Drug Sensitivity in Cancer (GDSC) resource and evaluated using patient-derived organoids (PDOs). RESULTS: Compared with OMBC, YMBC more frequently had human epidermal growth factor receptor 2 (HER2)-positive status (14.91% vs. 4.02%) and triple-negative tumors (4.92% vs. 1.78%), and had worse 5-year recurrence-free survival (hazard ratio=2.19, P=0.018). Transcriptomic analyses indicated enrichment of neural-related programs and reduced immune-related signaling in YMBC, and xCell/ESTIMATE supported lower immune infiltration. Consistently, WGCNA identified age-associated modules linking neural-related programs with reduced immune infiltration. Immunohistochemistry supported increased perineural invasion and lower CD8+ T cell infiltration in YMBC. GDSC-guided prioritization with PDO testing nominated sepantronium bromide (YM155) as a candidate vulnerability in YMBC. WES showed a higher NBPF10 mutation frequency in YMBC (54.5% vs. 14.3%, P<0.05). CONCLUSIONS: Integrated multi-omics profiling revealed age-stratified clinical and molecular heterogeneity in MBC. YMBC patients demonstrated inferior recurrence-free survival, neural signaling enrichment, an immune-cold microenvironment, and enriched NBPF10 mutations. These findings support age as a meaningful stratification variable in MBC risk assessment and treatment planning, and highlight the need for caution when considering treatment de-escalation in younger patients, while nominating YM155 as a candidate agent for prospective evaluation.

Male breast cancer↗

Integrated single-cell and spatial transcriptomic analyses reveal malignant epithelial glycolytic heterogeneity and spatial niche remodeling during colorectal cancer progression.

Colorectal cancer (CRC) progression is shaped by metabolic reprogramming and complex interactions within the tumor microenvironment. However, the cellular heterogeneity, spatial organization, and clinical relevance of glycolytic activity in CRC remain incompletely understood. In this study, we integrated single-cell RNA sequencing, bulk transcriptomics, and spatial transcriptomics data to systematically characterize glycolytic heterogeneity in CRC. Glycolytic activity was quantified using five independent scoring methods, consistently showing that epithelial cells exhibited the highest glycolytic activity across the two single-cell cohorts. Stratification of CopyKAT-verified aneuploid malignant epithelial cells into high-glycolysis (HG) and low-glycolysis (LG) subgroups by glycolysis scores revealed that HG cells exhibited higher stemness scores and chromosomal copy number variations. Cell-cell communication analysis revealed that, compared with LG cells, HG cells exhibited increased interaction frequency and strength with immune and stromal populations, indicating enhanced malignant epithelial-microenvironment crosstalk. Spatial transcriptomics analyses further revealed that glycolytic activity varied across normal colorectal tissue, primary CRC, and colorectal liver metastases, accompanied by progressive remodeling of epithelial-associated spatial niches and MIF-mediated intercellular communication. Bulk transcriptomic analysis identified a glycolysis-related prognostic signature with robust predictive performance, which served as an independent prognostic factor for overall survival in CRC cohorts. Collectively, these findings indicate that glycolytic heterogeneity is a key feature of CRC malignant epithelial cells and is closely associated with tumor progression, microenvironmental remodeling, and clinical outcomes.

Humans↗

Multi-omics integrative analysis provides insight into potential molecular responses to sustained high water flow in common carp (Cyprinus carpio) cultured in recirculating aquaculture.

To investigate the potential molecular responses by which water flow intensity affects the growth of common carp (Cyprinus carpio) in a recirculating aquaculture system (RAS), a control group (CG, actual water velocity 0.3&#xa0;cm/s) and three sustained flow treatment groups were established, including a low-flow group (LF, 1 body length per second, bl/s), a medium-flow group (MF, 2 bl/s), and a high-flow group (HF, 3 bl/s). After 12&#xa0;weeks of culture in the RAS, growth performance was compared among groups under different flow intensities. The best-performing group and the control group were then selected for the determination of intestinal digestive enzyme activities, as well as transcriptomic and whole-genome bisulfite sequencing analyses of muscle tissue. The results showed that the specific growth rate and feed intake of the HF group were significantly higher than those of the other groups (P&#xa0;<&#xa0;0.05), whereas no significant difference in feed conversion ratio was observed among groups. Compared with the CG group, lipase activity was significantly higher in the HF group (P&#xa0;<&#xa0;0.05), while &#x3b1;-amylase and trypsin activities showed increasing trends without significant differences. RNA-seq identified a total of 273 differentially expressed genes, including 72 upregulated genes and 201 downregulated genes in the HF group relative to the CG group. These genes were mainly enriched in glycolysis, pyruvate metabolism, ATP metabolism, the pentose phosphate pathway, the insulin signaling pathway, the PPAR signaling pathway, and the adipocytokine signaling pathway, indicating that sustained high water flow induced a muscle transcriptional response characterized by remodeling of energy metabolism and substrate utilization. Whole-genome bisulfite sequencing analysis showed that DNA methylation in common carp muscle occurred predominantly in the CpG context. Differentially methylated regions between the HF and CG groups were mainly distributed in transcription-related regulatory regions, including promoters, CpG islands, and CpG island shores. In promoter regions, the number of hypermethylated regions in the HF group relative to the CG group was markedly higher than that of hypomethylated regions. Integrated analysis further identified two candidate genes showing both promoter differential methylation and differential expression, namely LOC109094644 and bcorl1, suggesting that adaptation to high water flow may involve IGF-related growth regulation and remodeling of upstream transcriptional programs. The qPCR results were consistent with the transcriptomic data. Taken together, within the tested range, a sustained water flow of 3 bl/s was more conducive to the growth of common carp in the RAS, which may be associated with enhanced lipid digestion and utilization, remodeling of the muscle energy metabolic network, changes in promoter methylation, and the coordinated regulation of key candidate genes. This study provides a theoretical basis for clarifying the exercise adaptation mechanism of common carp in recirculating aquaculture and for optimizing flow velocity parameters.

Animals↗

A regulatory network underlying idiopathic pulmonary fibrosis.

BACKGROUND: Idiopathic pulmonary fibrosis (IPF) is a progressive interstitial lung disease in which genetic susceptibility interacts with epithelial, immune, and mesenchymal remodeling. Although the chromosome 11p15.5 locus contains established IPF susceptibility signals near MUC5B and TOLLIP, the broader regulatory architecture of this region remains incompletely resolved. METHODS: We integrated IPF genome-wide association study summary statistics with methylation, expression, and protein quantitative trait loci using summary-data-based Mendelian randomization (SMR). SMR-prioritized candidates were evaluated in independent transcriptomic and methylation cohorts and further contextualized using microRNA, transcription-factor, protein-interaction, machine-learning, single-cell, and spatial transcriptomic analyses. Fibrosis-associated expression patterns were assessed in a bleomycin-induced pulmonary fibrosis rat model. RESULTS: The analyses recovered the established MUC5B and TOLLIP signals and prioritized BRSK2 as a comparatively underexplored candidate supported by eQTL-based SMR and independent molecular evidence. The BRSK2 pQTL association did not pass the HEIDI test and was therefore not interpreted as convergent protein-level genetic evidence. Network analyses linked BRSK2 to cell-cycle, metabolic-stress, and senescence-related programs, while cross-cohort machine learning prioritized FOXA2, CDC25B, and NFE2 as informative network features. Single-cell and spatial analyses localized BRSK2 preferentially to fibroblast and myofibroblast compartments and to regions with greater histological fibrosis severity. In fibrotic rat lungs, BRSK2 expression increased, whereas FOXA2 and CDC25B decreased at the transcript and protein levels. CONCLUSIONS: These findings refine the molecular landscape of the chromosome 11p15.5 IPF susceptibility locus and prioritize BRSK2 as a candidate component of an IPF-associated profibrotic fibroblast state. Its causal contribution, direct regulatory relationships, and therapeutic tractability require targeted mechanistic validation.

Idiopathic Pulmonary Fibrosis↗

Multi-omics analysis identifies key genes and functional loci affecting teat number in American Large White and Landrace pigs and their application in optimizing genomic selection models.

BACKGROUND: Teat number is a crucial economic trait in pigs. It directly affects the ability of sows to lactate, which in turn influences the survival and health of piglets. The teat number of French Large White pigs is close to 16, while the teat number of American Large White and Landrace pigs is about 14. In order to improve the teat number of American Landrace and Large White pigs through molecular approaches and precise breeding techniques, we genotyped 2,131 American Landrace and 4,564 American Large White with teat number phenotype using a 50&#xa0;K SNP chip. Then, the SNP-chip data was imputed to the level of whole-genome sequencing (iWGS). Based on iWGS data, we conducted GWAS to identify novel, significant SNPs associated with teat number and to incorporate them into genomic selection. RESULTS: In Landrace pigs, significant SNPs for TTN mapped to SSC2, SSC7, SSC8, and SSC14; the SSC8 and SSC14 effects are novel. LTN mapped to SSC7, RTN to SSC7 and SSC8. The lead SSC7 SNP explained 2.60% of TTN phenotypic variance. In Large White pigs, significant SNPs were detected on SSC7 and SSC10 for TTN; SSC7, SSC10, and SSC12 for LTN; and SSC7 and SSC10 for RTN. The most significant locus on SSC7 accounted for 2.99% of the phenotypic variance in TTN. Additionally, a multi-population meta-analysis detected significant novel SNPs for LTN on SSC1 and SSC8. By utilizing Bayesian fine mapping, the most precise QTL confidence interval on SSC7 for both TTN and RTN in Large White pigs was reduced to 40&#xa0;kb. By integrating functional gene annotation with RNA-seq and ATAC-seq data from Erhualian and Bamaxiang pigs mammary placodes at embryonic day 26, we prioritized PTPN13, TRPV3, ZDHHC13, and BRD2 as novel candidate genes for teat number. We then incorporated the significant SNPs to GBLUP and benchmarked genomic-selection accuracy. In both breeds, fitting the top SNP as fixed maximized prediction for TTN and RTN, whereas treating all significant loci as an additional random effect optimized LTN. CONCLUSIONS: Our findings provide a theoretical basis for dissecting new key genes affecting teat number and for advancing molecular breeding of teat number in pigs.

Animals↗

Multi-omics Mendelian Randomization Prioritizes Neutrophil Extracellular Trap-related Genes Associated with Atrial Fibrillation Risk.

BACKGROUND: Neutrophil extracellular traps (NETs) participate in thrombosis, inflammation, and cardiovascular remodeling, yet whether NET-related genes (NRGs) are associated with atrial fibrillation (AF) risk across multiple molecular layers remains unclear. This study used a multiomics Mendelian randomization framework to prioritize NRGs supported by methylation, expression, and protein quantitative trait loci (QTL) data. METHODS: Genome-wide significant cis instruments (P < 5 &#xd7; 10-8) were obtained for 90 methylation QTLs (mQTLs), 100 expression QTLs (eQTLs), and 38 protein QTLs (pQTLs) mapped to 137 literature- curated NRG entries. Summary-data-based Mendelian randomization (SMR) coupled with the heterogeneity in dependent instruments (HEIDI) test was applied using whole-blood mQTL data (n = 1,980), eQTLGen blood eQTL data (n = 31,684), and deCODE plasma pQTL data (n = 35,559). AF outcome data were obtained from a meta-analysis including 60,620 cases and 970,216 controls of European ancestry. RESULTS: At the methylation level, 21 CpG-feature associations across 13 genes remained significant after HEIDI filtering and false discovery rate (FDR) correction. Expression-level analysis identified eight significant gene-AF associations, whereas protein-level analysis identified seven significant features representing five unique proteins. Cross-omics integration prioritized C3, MAPK3, and STAT3 as Tier 1 genes, CTSC, LPAR3, and THBD as Tier 2 genes, and fourteen additional genes as Tier 3 candidates. C3 showed risk-increasing protein-level associations together with multiple significant CpG signals, whereas MAPK3 and STAT3 showed directionally protective expression/protein or methylation/protein patterns. DISCUSSION: The cross-omics convergence on C3, MAPK3, and STAT3 is consistent with complement activation, immune-fibrotic signaling, and cytokine-regulatory pathways implicated in AF biology, but the findings should be interpreted as genetic prioritization rather than definitive intervention-ready causality. CpG-level heterogeneity at the C3 locus and the blood/plasma origin of the QTL resources further support a cautious interpretation. Modest colocalization support and the unresolved possibility of pQTL sample overlap further support this cautious, hypothesis-generating interpretation. CONCLUSION: Multi-omics SMR prioritizes C3, MAPK3, and STAT3 as the most consistently supported NET-related genes associated with AF risk. These findings provide a framework for atrialtissue replication and mechanistic validation of NET-related pathways in AF.

Atrial fibrillation↗

Machine learning-based clinical prediction model and multi-omics integration for assessing pancreatic cancer risk in new-onset diabetes.

BACKGROUND: Given that pancreatic cancer (PC) is typically diagnosed at an advanced stage but is often preceded by new-onset diabetes mellitus (NODM), providing a window for early detection, we sought to develop and validate an interpretable machine-learning model integrated with multi-omics profiling to identify early biomarkers of NODM-associated PC. METHODS: In a population-based cohort, individuals with NODM-associated PC and NODM without PC were identified and randomly divided (70:30) into training and validation sets after feature selection. Eight machine learning (ML) classifiers were compared using fivefold cross-validation, and model performance was evaluated in terms of discrimination, calibration, and decision curve&#x2013;based clinical utility. We evaluated interpretability using the Shapley additive explanations (SHAP) analyses. Mechanistically, Olink proteomic profiling and metabolomics were analyzed through clinical classifications and model-defined risk strata. RESULTS: Categorical boosting achieved the best performance in the independent validation set (AUROC&#x2009;=&#x2009;0.844). The NODM cohort was stratified into high- (n&#x2009;=&#x2009;2,362) and low-risk (n&#x2009;=&#x2009;5,030) groups, and internal validation together with SHAP analyses demonstrated consistent model performance and identified clinically interpretable predictors. Proteomic and metabolomic analyses under clinical and risk-based grouping identified 39 overlapping differentially expressed proteins and 145 overlapping metabolites with enriched across 11 shared KEGG pathways. Cross-platform validation highlighted PLTP, CRTAC1, and ITGAV as serum biomarkers with a strong potential for early NODM-PC detection. CONCLUSIONS: We developed an interpretable ML framework centered on NODM enables practical risk stratification for early PC detection by multi-omics and provides a pathway of ML-based triage followed by biomarker confirmation for earlier detection and diagnosis.

Humans↗

Multi-Omics Integration Identifies a Five-Gene Metabolic Signature With Experimental Validation in Clear Cell Renal Cell Carcinoma.

BACKGROUND: Clear cell renal cell carcinoma (ccRCC) is hallmarked by profound metabolic reprogramming; however, its intricate crosstalk with the tumor immune microenvironment (TIME) and its clinical ramifications remain inadequately elucidated. This study aims to systematically decipher the metabolic-immune interplay in ccRCC through multi-omics integration, with the goal of identifying robust prognostic biomarkers and actionable therapeutic vulnerabilities. AIMS: This study aims to systematically decipher the metabolic-immune interplay in clear cell renal cell carcinoma (ccRCC) through multi&#x2011;omics integration, and to identify robust prognostic biomarkers and actionable therapeutic vulnerabilities that can inform precision risk stratification and individualized treatment strategies. METHODS: We integrated bulk transcriptomic, genomic, and clinical data from multiple ccRCC cohorts. Differential expression and functional enrichment analyses were performed to characterize metabolic pathway alterations. Mendelian randomization (MR) was employed to infer causal relationships between metabolic disorders and ccRCC risk. A machine learning-based prognostic framework, incorporating SHAP (SHapley Additive exPlanations) for feature interpretability, was constructed and rigorously validated. TIME heterogeneity was dissected using deconvolution algorithms, while drug sensitivity, tumor mutation burden (TMB), and TIDE scores were utilized to assess therapeutic responses and immune evasion. Candidate gene function was evaluated through in&#xa0;vitro gain- and loss-of-function assays, with expression validated via TCGA, HPA, western blot, and qRT-PCR. RESULTS: Enrichment analysis identified coordinated dysregulation in lipid metabolism, energy homeostasis, and hypoxia response pathways. MR analysis confirmed lipid metabolism disorders as a causal risk factor for ccRCC. Our machine-learning model, centered on five core SHAP-identified features (SUCLA2, ACAT1, PC, SUCLG1, and HMGCS2), demonstrated superior predictive accuracy over conventional clinical staging. Immune profiling unveiled dichotomous TIME states: the low-risk group retained active immune surveillance, whereas the high-risk group was enriched with immunosuppressive subsets. Drug sensitivity screening pinpointed LY2109761 and carmustine as high-risk-specific candidate agents. Furthermore, TMB and TIDE analyses stratified high-risk patients displaying genomic instability and immune evasion phenotypes. Functionally, SUCLA2 knockdown significantly enhanced ccRCC cell proliferation and invasion, while its overexpression suppressed these malignant phenotypes, corroborating its tumor-suppressive role. Expression patterns of the hub genes were consistently validated across multi-level datasets and experimental assays. CONCLUSION: This study establishes a precision oncology framework for ccRCC by functionally linking metabolic biomarkers, immunophenotypes, and stratified therapeutic strategies. Importantly, we identify SUCLA2 as a potential functional tumor suppressor and a promising target for further mechanistic and translational investigation.

Humans↗

Multi-omic biomarkers in cardiovascular disease: Discovery to clinical translation.

Cardiovascular disease (CVD) remains the leading cause of mortality worldwide, necessitating improved risk stratification and early detection strategies. Multiomics approaches that integrate genomics, transcriptomics, proteomics, metabolomics, and epigenomics offer unprecedented opportunities for biomarker discovery and precision medicine in cardiovascular care. This narrative review examines the current landscape of multiomics biomarkers for CVD, tracing their evolution from discovery to clinical translation. We synthesize evidence from recent studies evaluating the clinical utility of integrated omics approaches across diverse cardiovascular conditions, including atherosclerotic cardiovascular disease, heart failure, and atrial fibrillation. High-throughput proteomics has identified novel protein signatures that enhance cardiovascular risk prediction beyond traditional risk factors. Metabolomics has revealed pathway-specific biomarkers, including trimethylamine N-oxide and lipid species, associated with atherogenesis. Polygenic risk scores derived from genomic data demonstrate incremental value when combined with clinical risk scores. Multiomics biomarkers represent a transformative approach to cardiovascular risk assessment and disease management.

Humans↗

Multi-omics panorama of glaucoma: Pathogenesis, biomarkers, and novel therapeutic strategies.

Glaucoma is a group of irreversible, blinding eye diseases characterized by progressive loss of retinal ganglion cells, leading to gradual visual field defects that severely impact patients' quality of life. Its complex pathophysiological mechanisms remain incompletely understood, limiting the development of early diagnostic and effective therapeutic strategies. Advances in omics technologies have provided new insights into elucidating the pathophysiology of glaucoma. We summarize specific alterations in genomics, transcriptomics, proteomics, metabolomics, epigenomics, and microbiomics associated with glaucoma. We emphasize the systematic analysis of disease mechanisms, identification of clinically applicable biomarkers, and discovery of novel therapeutic targets through the integration of these data. This approach paves new pathways for glaucoma subtype diagnosis and personalized treatment, while also outlining future research directions and challenges.

Humans↗

Decoding the molecular basis of blue grain color codominance in Qingke: Integrative analysis of RNA-seq, DNA methylation, and miRNA-seq.

The grains on single spike of the F1 generation from the cross between blue- and white-grained Qingke (Hordeum vulgare L. var. nudum Hook. f.) are randomly distributed in blue and white colors. This study integrated data from RNA-seq, DNA methylation, and miRNA-seq to analyze this trait. The results showed that the HvF3'5'H gene is likely central to the development of this codominant phenotype. Through cross-validation of three omics approaches, it was found that the HvMYB gene targeted by miR858-z, as well as the WRKY24 and At3g44326 genes targeted by novel-m0152-5p, novel-m0153-5p, and novel-m0154-5p, are correlated with DNA methylation. qRT-PCR analysis confirmed that the four aforementioned genes exhibited variety-specific and developmental stage-specific expression patterns. This study dissects the regulatory network underlying the codominant blue and white grain color divergence on a single Qingke spike from a multi-omics perspective.

DNA Methylation↗

HoloFoodR: a statistical programming framework for holo-omics data integration workflows.

SUMMARY: Holo-omics is an emerging research area that integrates multi-omic datasets from the host organism and its microbiome to study their interactions. Recently, curated and openly accessible holo-omic databases have been developed. The HoloFood database, for instance, provides nearly 10 000 holo-omic profiles for salmon and chicken under controlled treatments. However, bridging the gap between holo-omic data resources and algorithmic frameworks remains a challenge. Combining the latest advances in statistical programming with curated holo-omic data sets can facilitate the design of open and reproducible research workflows in the emerging field of holo-omics. AVAILABILITY AND IMPLEMENTATION: HoloFoodR R/Bioconductor package and the source code are available under the open-source Artistic License 2.0 at the package homepage https://doi.org/10.18129/B9.bioc.HoloFoodR.

Software↗

Radiogenomics predicts immune microenvironment heterogeneity and response to combination immunotherapy in hepatocellular carcinoma.

BACKGROUND: The combination of immune checkpoint inhibitors (ICIs) with anti-angiogenic agents is the preferred first-line therapy option for patients with advanced hepatocellular carcinoma (HCC), yet only a subset of patients responds, urging the quest for prediction biomarkers. We aimed to integrate genomics with radiology to propose an immune-derived radiogenomics biomarker of response to such combination immunotherapy and evaluate its added value in clinical context. METHODS: We integrated bulk RNA sequencing (RNA-seq) and proteomics data of 994 HCC patients with single-cell RNA-seq data of 11 samples across multiple datasets to identify an immune-related signature (IRS) that may influence sensitivity or resistance to such combined immunotherapy strategy, followed by verification of selected marker genes using immunohistochemistry and cytological experiments. We then trained/validated a cross-modality radiogenomics biomarker using machine learning based on TCIA database that was further tested in multi-scale independent cohorts covering 754 HCC patients. RESULTS: Integrative multi-omics analysis identifed a parsimonious 2-gene prognostic signature including KPNA2 and SMG5 that was significantly associated with immune heterogeneity and response to combination immunotherapy. Machine-learning pipeline exported the optimal 4-feature radiogenomics biomarker using support vector machine that significantly discriminated prognosis (hazard ratio 1.415&#x2013;1.890; p&#x2009;<&#x2009;0.05 for all) and modestly predicted response to ICI plus anti-angiogenic therapy (area under the curve 0.720&#x2013;0.829) in independent retrospective series across major imaging modalities (computed tomography/magnetic resonance imaging). In a prospective neoadjuvant cohort, this biomarker also showed favorable performance for predicting pathological response and tumor recurrence, accompanied by biological validation through single-cell RNA-seq analysis of pre-treatment biopsies. CONCLUSIONS: Our study provides a cross-device-cross-modal radiogenomics biomarker that can improve patient selection for emerging ICI plus anti-angiogenic therapy with novel potential therapeutic targets in HCC.

Humans↗