Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “transcriptome analysis”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 577 records · Page 32Linked to original sources

Current and future applications of SAGE to cardiovascular medicine.

The recently sequenced mammalian genomes represent unprecedented resources for advancing our understanding of human diseases. Characterizing gene expression is an important step in translating genomic sequences into clinically useful information. Currently, gene expression studies are revolutionizing the approaches taken to address both basic science and clinical questions. Two major methods have emerged for the global examination of the transcriptome: microarrays and Serial Analysis of Gene Expression (SAGE). The SAGE technique comprehensively maps gene transcription by using the genomic database, yet it remains relatively underutilized for studying cardiovascular biology. This review describes current cardiovascular studies using the SAGE technique and outlines some potential strategies for employing this powerful tool to further our understanding of the cardiovascular system in health and in disease.

Animals↗

Heterozygous knockout of Synaptotagmin13 phenocopies ALS features and TP53 activation in human motor neurons.

Spinal motor neurons (MNs) represent a highly vulnerable cellular population, which is affected in fatal neurodegenerative diseases such as amyotrophic lateral sclerosis (ALS) and spinal muscular atrophy (SMA). In this study, we show that the heterozygous loss of SYT13 is sufficient to trigger a neurodegenerative phenotype resembling those observed in ALS and SMA. SYT13+/- hiPSC-derived MNs displayed a progressive manifestation of typical neurodegenerative hallmarks such as loss of synaptic contacts and accumulation of aberrant aggregates. Moreover, analysis of the SYT13+/- transcriptome revealed a significant impairment in biological mechanisms involved in motoneuron specification and spinal cord differentiation. This transcriptional portrait also strikingly correlated with ALS signatures, displaying a significant convergence toward the expression of pro-apoptotic and pro-inflammatory genes, which are controlled by the transcription factor TP53. Our data show for the first time that the heterozygous loss of a single member of the synaptotagmin family, SYT13, is sufficient to trigger a series of abnormal alterations leading to MN sufferance, thus revealing novel insights into the selective vulnerability of this cell population.

Humans↗

The offonome reveals on and off states of gene expression near the detection limit of RNA-seq.

RNA-seq, widely used for gene expression profiling, provides nucleotide level genome coverage and summary gene expression values. Generally, low-expressed genes are ignored due to their unfavorable signal-to-noise ratio, however, these genes may offer crucial information, such as detecting rare cells in bulk tissues. In this study, we applied an approach that transforms the expression levels of low-expressed genes into a robust dichotomized on/off state by leveraging similarities in transcript coverage shape. Applied to three human cancer cohorts from the Cancer Genome Atlas (TCGA), chosen based on tissue morphology and anatomic site, we identified genes, the "offonome" near the detection limit, consistently or occasionally off across samples. Genes in the offonome spectrum proved useful for supervised and unsupervised applications, including characterizing oncogenic pathways, and identifying rare populations of cells in bulk tissue. Interrogating the offonome is relevant to bulk tumor analyses like TCGA, potentially expediting gene investigation in low-input situations like single cell RNA-seq.

Humans↗

Long noncoding RNA LIRIL2R modulates FOXP3 levels and suppressive function of human CD4+ regulatory T cells by regulating IL2RA.

Regulatory T cells (Tregs) are central in controlling immune responses, and dysregulation of their function can lead to autoimmune disorders or cancer. Despite extensive studies on Tregs, the basis of epigenetic regulation of human Treg development and function is incompletely understood. Long intergenic noncoding RNAs (lincRNA)s are important for shaping and maintaining the epigenetic landscape in different cell types. In this study, we identified a gene on the chromosome 6p25.3 locus, encoding a lincRNA, that was up-regulated during early differentiation of human Tregs. The lincRNA regulated the expression of interleukin-2 receptor alpha (IL2RA), and we named it the lincRNA regulator of IL2RA (LIRIL2R). Through transcriptomics, epigenomics, and proteomics analysis of LIRIL2R-deficient Tregs, coupled with global profiling of LIRIL2R binding sites using chromatin isolation by RNA purification, followed by sequencing, we identified IL2RA as a target of LIRIL2R. This nuclear lincRNA binds upstream of the IL2RA locus and regulates its epigenetic landscape and transcription. CRISPR-mediated deletion of the LIRIL2R-bound region at the IL2RA locus resulted in reduced IL2RA expression. Notably, LIRIL2R deficiency led to reduced expression of Treg-signature genes (e.g., FOXP3, CTLA4, and PDCD1), upregulation of genes associated with effector T cells (e.g., SATB1 and GATA3), and loss of Treg-mediated suppression.

Humans↗

An ancestral secretory apparatus in the protozoan parasite Giardia intestinalis.

The protozoan parasite Giardia intestinalis belongs to one of the earliest diverged eukaryotic lineages. This is also reflected in a simple intracellular organization, as Giardia lacks common subcellular compartments such as mitochondria, peroxisomes, and apparently also a Golgi apparatus. During encystation, developmentally regulated formation of large secretory compartments containing cyst wall material occurs. Despite the lack of any morphological similarities, these encystation-specific vesicles (ESVs) show several biochemical characteristics of maturing Golgi cisternae. Previous studies suggested that Golgi structure and function are induced only during encystation in Giardia, giving rise to the hypothesis that ESVs, as a Giardia Golgi equivalent, are generated de novo. Alternatively, ESV compartments could be built on the template structure of a cryptic Golgi in trophozoites in response to ER export of cyst wall material during encystation. We addressed this question by defining the molecular framework of the Giardia secretory apparatus using a comparative genomic approach. Analysis of the corresponding transcriptome during growth and encystation revealed surprisingly little stage-specific regulation. A panel of antibodies was generated against selected marker proteins to investigate the developmental dynamics of the endomembrane system. We show evidence that Giardia accommodates the export of large amounts of cyst wall material through re-organization of membrane compartment(s) in trophozoites with biochemical similarities to ESVs. This suggests that ESVs are selectively stabilized Golgi-like compartments in a unique and archetypical secretory system, which arise from a structural template in trophozoites rather than being generated de novo.

Animals↗

Identification and analysis of chromodomain-containing proteins encoded in the mouse transcriptome.

The chromodomain is 40-50 amino acids in length and is conserved in a wide range of chromatic and regulatory proteins involved in chromatin remodeling. Chromodomain-containing proteins can be classified into families based on their broader characteristics, in particular the presence of other types of domains, and which correlate with different subclasses of the chromodomains themselves. Hidden Markov model (HMM)-generated profiles of different subclasses of chromodomains were used here to identify sequences encoding chromodomain-containing proteins in the mouse transcriptome and genome. A total of 36 different loci encoding proteins containing chromodomains, including 17 novel loci, were identified. Six of these loci (including three apparent pseudogenes, a novel HP1 ortholog, and two novel Msl-3 transcription factor-like proteins) are not present in the human genome, whereas the human genome contains four loci (two CDY orthologs and two apparent CDY pseudogenes) that are not present in mouse. A number of these loci exhibit alternative splicing to produce different isoforms, including 43 novel variants, some of which lack the chromodomain. The likely functions of these proteins are discussed in relation to the known functions of other chromodomain-containing proteins within the same family.

Acetyltransferases↗

Long-read sequencing reveals widespread novel splicing and neojunction-derived neoantigens in nasopharyngeal carcinoma.

The widespread transcriptomic diversity driven by alternative splicing (AS) contributes to all hallmarks of cancer and represents a critical source of neoantigens for personalized immunotherapy. However, unlike other major malignancies, the full repertoire of AS in nasopharyngeal carcinoma (NPC) remains underexplored. Here, we employ long-read sequencing (LR-seq) to generate a high-resolution, isoform-level transcriptomic atlas from a cohort of 14 NPC tumor samples and four immortalized nasopharyngeal epithelial cell lines. We identify a substantial number of full-length novel transcripts (22,687; ∼44.38%), which reveal diverse splicing patterns and previously unannotated splicing events. By integrating short-read RNA-seq data to quantify isoform expression, we discover a subset of novel transcripts that are differentially expressed between tumor samples and immortalized nasopharyngeal epithelial cell lines. Furthermore, LR-seq enables precise identification of chimeric readthrough fusion transcripts, such as CLDN15-FIS1 and FOXRED2-TXN2 Finally, we develop a computational framework, tumor-specific splicing neoantigen detection (TS-SNAD), to predict neoantigens originating from novel exon-exon junctions (neojunctions) in tumor-specific novel transcripts. Using this framework, we identify neojunction-derived neoantigens and experimentally validate the immunogenicity of selected HLA-B*40:01-restricted neoantigens. These neojunction-derived peptides constitute a new class of noncanonical neoantigens with significant potential for developing personalized cancer vaccines for NPC.

Humans↗

Unveiling m7G modification patterns and causal drivers governing intracranial aneurysm rupture risk through multi-omics validation and m7G-MeRIP-seq profiling.

Intracranial aneurysm (IA) rupture causes severe brain hemorrhage with high mortality, yet its molecular drivers remain unclear and better risk prediction is urgently needed. Using transcriptomics, single-cell analysis, and genetic data, we investigated the role of N7-methylguanosine (m7G) RNA modification in IA. We identified distinct m7G modification patterns, validated their methylation features in patient samples, and incorporated these patterns into a machine learning-based rupture prediction model. The presence and characteristics of m7G patterns significantly improved model performance, achieving high predictive accuracy across three independent cohorts (AUC 0.91-0.95). Genetic analyses further identified three causal m7G-related genes (NSUN2, IFIT5, SNUPN), and laboratory experiments confirmed their altered expression and methylation in ruptured aneurysms. Overall, our findings demonstrate that m7G modifications play a key role in IA rupture. The validated prediction model offers strong clinical potential for rupture risk assessment, and the identified genes represent promising therapeutic targets.

Humans↗

Characterization of METTL3/14-mediated m6A modification in human transcriptome using Nanopore direct RNA sequencing.

Post-transcriptional RNA modifications modulate diverse aspects of RNA metabolism. N6-methyladenosine (m6A), one of the most abundant internal RNA modifications, is deposited by the core methyltransferase complex, METTL3 and METTL14. Oxford Nanopore Technologies (ONT) platform permits direct, single RNA molecule sequencing while preserving native modifications. However, without rigorous benchmarking, the accuracy and reproducibility of modification detection remain uncertain. Here, we leveraged ONT to comprehensively profile bona fide m6A modifications in cellular RNAs at single-nucleotide resolution by integrating two direct RNA sequencing chemistries (RNA002 and RNA004) with the m6Anet and Dorado modification-detection models. We independently depleted METTL3 and METTL14 in human cells and rigorously validated modification calls through several assays and independent orthogonal methods (GLORI and miCLIP). We find that Dorado detected a higher number of m6A events and enabled simultaneous detection of other RNA modifications (5-methylcytosine, pseudouridine, and inosine). Pairing Dorado with an in vitro transcribed, unmodified control under stringent filtering, we provide compelling evidence supporting a global reduction in m6A sites and stoichiometry within coding sequences and across genes, particularly in highly modified genes and sites, and at consensus DRACH motifs. We report a differential and complex regulation of modified transcripts, accompanied by a global reduction in poly(A) tail length. Notably, METTL3 and METTL14 depletion produced distinct transcript-specific effects, supporting non-redundant roles within the m6A writer complex. Together, our study illustrates a notable advancement of ONT capabilities and establishes a robust transcriptome-wide framework for RNA modification detection, thereby laying the groundwork for exploring the contribution of METTL3/METTL14 to cellular functions and disease.

Humans↗

A Computational Workflow for Prioritizing Microbial Metabolite-Associated Host Genes in Constipation-Predominant Irritable Bowel Syndrome.

No standardized computational pipeline exists for systematically prioritizing microbial metabolite-associated host genes and protein-ligand complexes from publicly available chemical, genomic, and structural databases. This article describes an eight-stage workflow that accepts a user-defined set of gut microbiota-derived metabolites and produces a ranked shortlist of candidate metabolite-associated host genes, enriched biological pathways, and structurally prioritized protein-ligand complexes for experimental follow-up. The pipeline integrates (i) chemoinformatic metabolite profiling; (ii) multi-database candidate target prediction using protein-chemical interaction and ligand-based target-prediction tool and a molecular docking program; (iii) differential gene expression analysis of publicly available transcriptomic data; (iv) target-differentially expressed gene overlap; (v) protein-protein interaction network construction and pathway enrichment; (vi) molecular docking with a molecular docking program; (vii) 200 ns molecular dynamics simulation using a molecular dynamics engine with a protein force field used for molecular dynamics simulations; and (viii) MM-PBSA binding free-energy estimation. As a worked example, nine gut microbiota-derived or microbiota-modified metabolites representing short-chain fatty acids, bile acids, tryptophan-derived metabolites, and urolithin A were processed using the public IBS-C rectal mucosal transcriptomic dataset GSE36701. The workflow ranked 17 unique predicted metabolite-associated genes that were differentially expressed in this dataset. Docking, molecular dynamics simulation, and MM-PBSA analyses structurally prioritized five metabolite-protein complexes: lithocholic acid-VDR, lithocholic acid-NR1H4/FXR, ursodeoxycholic acid-NR1H4/FXR, tryptamine-HTR2A (simulated in an explicit 1-Palmitoyl-2-oleoyl-sn-glycero-3-phosphocholine (POPC) lipid bilayer), and urolithin A-CASP3. The protocol is designed to be adaptable to other metabolite sets, disease transcriptomic datasets, and target classes; all outputs are hypothesis-generating computational predictions that require independent transcriptomic replication, protein-level validation, and functional ligand-response assays before causal or therapeutic conclusions can be drawn.

Irritable Bowel Syndrome↗

Amaranth: enhanced single-cell transcript assembly via discriminative modelling of UMI reads and internal reads.

MOTIVATION: Single-cell RNA sequencing (scRNA-seq) has transformed transcriptome profiling at cellular resolution, yet accurate reconstruction of full-length transcripts for individual cells remains a central challenge. Emerging scRNA-seq protocols can produce reads that span entire transcripts, enabling isoform-level expression analysis. For example, Smart-seq protocols combine unique molecular identifier (UMI)-linked reads that index and stitch together multiple reads from the same molecule, with internal reads filling coverage gaps. We demonstrate that these read types exhibit markedly different biological and statistical properties in strandness, 5'/3' coverage bias, and genomic locality. Existing assemblers fail to leverage these distinctions, yielding suboptimal assembly. RESULTS: We developed Amaranth, a novel single-cell assembler that discriminatively models UMI and internal reads. Amaranth implements heuristics specifically designed to address the distinct biases of UMI-linked and internal reads, enabling accurate strandness assignment for internal reads, reliable splicing graph refinement, and precise transcript start site determination. We also developed Amaranth-meta, which integrates information across cells to enhance individual cell assemblies. Benchmarked on Smart-seq3 datasets from human HEK293T and mouse fibroblast cells, Amaranth outperformed other state-of-the-art assemblers in assembling individual cells and in meta-assembly. Amaranth advances isoform-level analysis in single-cell transcriptomics, facilitating detailed studies at cellular resolution. AVAILABILITY AND IMPLEMENTATION: Amaranth is implemented in C++ and is freely available at https://github.com/Shao-Group/amaranth under the BSD-3-Clause license. Scripts, documentation, and data for reproducing experiments in this manuscript are available at https://github.com/Shao-Group/amaranth-test.

Single-Cell Gene Expression Analysis↗

Transcriptome characterization of the dimorphic and pathogenic fungus Paracoccidioides brasiliensis by EST analysis.

Paracoccidioides brasiliensis is a pathogenic fungus that undergoes a temperature-dependent cell morphology change from mycelium (22 degrees C) to yeast (36 degrees C). It is assumed that this morphological transition correlates with the infection of the human host. Our goal was to identify genes expressed in the mycelium (M) and yeast (Y) forms by EST sequencing in order to generate a partial map of the fungus transcriptome. Individual EST sequences were clustered by the CAP3 program and annotated using Blastx similarity analysis and InterPro Scan. Three different databases, GenBank nr, COG (clusters of orthologous groups) and GO (gene ontology) were used for annotation. A total of 3,938 (Y = 1,654 and M = 2,274) ESTs were sequenced and clustered into 597 contigs and 1,563 singlets, making up a total of 2,160 genes, which possibly represent one-quarter of the complete gene repertoire in P. brasiliensis. From this total, 1,040 were successfully annotated and 894 could be classified in 18 functional COG categories as follows: cellular metabolism (44%); information storage and processing (25%); cellular processes-cell division, posttranslational modifications, among others (19%); and genes of unknown functions (12%). Computer analysis enabled us to identify some genes potentially involved in the dimorphic transition and drug resistance. Furthermore, computer subtraction analysis revealed several genes possibly expressed in stage-specific forms of P. brasiliensis. Further analysis of these genes may provide new insights into the pathology and differentiation of P. brasiliensis.

Base Sequence↗

A Molecularly Anchored Spatial Transcriptomic Framework for Precise CA1-Subiculum Parcellation and Region-Resolved Analysis in Alzheimer's Disease.

BACKGROUND: The precise molecular delineation of the interface between the Subiculum (Sub) and cornu ammonis 1 (CA1) is a challenge in hippocampal research, as conventional cytoarchitectural boundaries are often ambiguous and limit reproducible regional annotation. Here, we developed a molecularly anchored spatial transcriptomic framework to define CA1-Sub regional identities using high-definition spatial transcriptomics (Stereo-seq) and single-nucleus RNA sequencing (snRNA-seq) references. FINDINGS: Using a human hippocampal Stereo-seq dataset from 12 donors, we established a data-driven parcellation framework that defines reproducible molecular features distinguishing CA1 and Sub while capturing the transition between these regions. FN1 was identified as a Sub-enriched marker in a subset of EX_Sub and, together with ETV1 and additional regional markers, enabled molecular assignment of CA1 and Sub identities across datasets. The Sub association of FN1 and ETV1 was further supported by human 10X Genomics spatial transcriptomics, mouse in situ hybridization data, and a mouse spatial transcriptomic dataset. Applying this framework to Alzheimer's disease (AD) tissues revealed region-specific transcriptional alterations across CA1 and Sub, including enrichment of mitochondrial energy metabolism-related transcripts in the Sub, suggesting exploratory transcriptional associations of altered metabolic function. CONCLUSIONS: This study provides a molecularly anchored framework for human CA1-Sub parcellation that complements conventional annotation. By defining regional molecular states while preserving the biological continuum across CA1-Sub interface, this approach enables more consistent regional analysis of human hippocampus tissue across donors, datasets, and disease conditions.

Journal Article↗

Genome-wide identification of WOX transcription factors and functional characterization of WOX4 and WOX13 involved in cold stress response in Malus baccata.

INTRODUCTION: Cold stress is a major abiotic threat to apple production. Malus baccata has exceptional cold hardiness and is widely used as a superior cold-resistant rootstock. The WUSCHEL-related homeobox (WOX) transcription factor family regulates plant growth, development and stress adaptation, whereas the functions of WOX genes in cold tolerance of M. baccata remain elusive. METHODS: In the present work, 19 MbWOX family members were identified and characterized at the genome-wide level. Evolutionary analysis, cis-element prediction, transcriptome profiling and real-time quantitative PCR (RT-qPCR) were performed to screen core cold-responsive genes. Overexpression vectors were constructed and transformed into Arabidopsis seedlings for functional verification. RESULTS: Evolutionary analysis revealed that segmental duplication drove the expansion of the MbWOX family, and these genes contained a variety of stress-responsive cis-elements. Combined transcriptome and RT-qPCR analyses confirmed that MbWOX4 and MbWOX13 were core cold-responsive genes with distinct expression patterns. The two genes participated in cold signal transduction by interacting with different transcription factor networks. Functional tests revealed that MbWOX4 and MbWOX13 isoforms differentially modulated seedling cold tolerance under low-temperature stress.

Malus baccata↗

A pluripotent stem cell atlas of multilineage differentiation.

Human pluripotent stem cells offer a scalable platform to study genetic and signalling mechanisms governing cell lineage decisions during differentiation. Genome-wide and single-cell transcriptomics technologies likewise offer high-throughput analysis of heterogeneous cell differentiation states. While in vivo development has been extensively characterised using these technologies, there remains a need for comprehensive single-cell transcriptomic profiling of stem cell differentiation from pluripotency. Understanding gene expression changes governing differentiation in vitro is key to developing high fidelity differentiation protocols and understanding fundamental mechanisms of development. We generated a single-cell RNA sequencing time course to study the role of developmental signalling pathways on multilineage diversification from pluripotency in vitro. The combined dataset of over 60,000 cells spans cell types from a time course of differentiation across all germ layers, ranging from gastrulation cell states to progenitor and committed cell types. These data provide a diverse benchmarking reference point to compare against in vivo development and advance understanding of signalling regulation of differentiation, providing insights into protocol development, drug screening, and regenerative medicine applications.

Pluripotent Stem Cells↗

Mouse proteome analysis.

A general overview of the protein sequence set for the mouse transcriptome produced during the FANTOM2 sequencing project is presented here. We applied different algorithms to characterize protein sequences derived from a nonredundant representative protein set (RPS) and a variant protein set (VPS) of the mouse transcriptome. The functional characterization and assignment of Gene Ontology terms was done by analysis of the proteome using InterPro. The Superfamily database analyses gave a detailed structural classification according to SCOP and provide additional evidence for the functional characterization of the proteome data. The MDS database analysis revealed new domains which are not presented in existing protein domain databases. Thus the transcriptome gives us a unique source of data for the detection of new functional groups. The data obtained for the RPS and VPS sets facilitated the comparison of different patterns of protein expression. A comparison of other existing mouse and human protein sequence sets (e.g., the International Protein Index) demonstrates the common patterns in mammalian proteomes. The analysis of the membrane organization within the transcriptome of multiple eukaryotes provides valuable statistics about the distribution of secretory and transmembrane proteins

Animals↗

Interaction of soft condensed materials with living cells: phenotype/transcriptome correlations for the hydrophobic effect.

The assessment of biomaterial compatibility relies heavily on the analysis of macroscopic cellular responses to material interaction. However, new technologies have become available that permit a more profound understanding of the molecular basis of cell-biomaterial interaction. Here, both conventional phenotypic and contemporary transcriptomic (DNA microarray-based) analysis techniques were combined to examine the interaction of cells with a homologous series of copolymer films that subtly vary in terms of surface hydrophobicity. More specifically, we used differing combinations of N-isopropylacrylamide, which is presently used as an adaptive cell culture substrate, and the more hydrophobic, yet structurally similar, monomer N-tert-butylacrylamide. We show here that even discrete modifications with respect to the physiochemistry of soft amorphous materials can lead to significant impacts on the phenotype of interacting cells. Furthermore, we have elucidated putative links between phenotypic responses to cell-biomaterial interaction and global gene expression profile alterations. This case study indicates that high-throughput analysis of gene expression not only can greatly refine our knowledge of cell-biomaterial interaction, but also can yield novel biomarkers for potential use in biocompatibility assessment.

Cell Adhesion↗

Identification of candidate genes and regulatory mechanisms for thick shank phenotype of Dong Tao chickens.

The Dong Tao Chicken has garnered widespread attention owing to its hallmark trait of remarkably thick shanks. However, the genetic mechanisms and molecular basis underlying this unique phenotype remain largely elusive to date. We carried out crossbreeding trials, performed genome-wide association study (GWAS), and completed transcriptome sequencing of leg tissues. Crossbreeding experiments confirmed that the thick-leg trait of Dong Tao chickens is polygenically controlled. Furthermore, GWAS analysis on the F₂ segregating population screened NAKIN3 as a key candidate gene associated with this trait. Transcriptomic results indicated that tarsometatarsal skin acts as the primary tissue regulating shank circumference growth. Moreover, ACTB was recognized as a core gene driving dermal thickening of the tarsometatarsus. By interacting with IRF family members, TLR4, SOS1, IFNG, STAT family members, EGF and other molecules, ACTB participates in multiple Kyoto Encyclopedia of Genes and Genomes (KEGG) pathways including the Toll-like receptor signaling pathway and cytokine-cytokine receptor interaction pathway to coordinately modulate dermal thickening in the tarsometatarsal skin of Dong Tao chickens. This study identifies key loci and genes governing shank development, offering theoretical basis and genetic resources for dissecting molecular mechanisms and selective breeding of Dong Tao chickens. Nevertheless, these preliminary findings need to be further verified by expanding the experimental population and sample size, together with additional independent validation experiments.

Dong Tao chicken↗