Search PubMedSearch

SEARCH · Search PubMed

Results for “Aptamers, Nucleotide”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

33 records · Page 2Linked to original sources

Cloning of two Hsp70 genes and association analysis between SNP haplotypes and high temperature tolerance trait in red swamp crayfish (Procambarus clarkii).

Aquaculture is suffering the challenge from high temperature climate. Two Hsp70 genes, PcHsp70-1 and PcHsp70-2, as key genes involved in the high temperature tolerance of red swamp crayfish (Procambarus clarkii) were identified and cloned in this study. Their molecular features and expression patterns were characterized, revealing the distinct tissue-specific upregulation expression under high temperature stress (33 °C). Two SNPs, PcHsp70-1 (SNP258) and PcHsp70-2 (SNP555) were examined to associate with high temperature tolerance in three populations (n = 675). The genotypes of PcHsp70-1-SNP258 (GA) and PcHsp70-2-SNP555 (TT) were significantly associated with stronger high temperature tolerance. Notably, individuals carrying the haplotype of Hap I (GG + TT) showed a survival rate exceeding 70% under high temperature stress, whereas, the Hap VIII (AA + CT) showed it at 5.2%. RNA interference of PcHsp70-1 resulted in a significant decrease expression of the gene GSH-Px and its encoding protein (glutathione peroxidase) activity, and damage in intestinal tissue under high temperature stress. The transcriptome result revealed that PcHsp70-1 participates in regulation of the pathways related to cytoskeletal construction, immune response, apoptosis, and antioxidant defense. These findings indicate that PcHsp70 genes are crucial for the cellular stress response under high temperature stress. The developed Kompetitive Allele Specific PCR (KASP) markers provide valuable tools for the marker-assisted selection of high temperature tolerant crayfish varieties, supporting the sustainable development of aquaculture under the challenge of global warming.

Animals

Protective association of the ELMO1 rs741301 variant against diabetes mellitus and diabetic nephropathy: a systematic meta-analysis of case-control studies.

CONTEXT: Diabetes mellitus (DM), an endocrine disorder, is characterised by persistently elevated blood glucose levels due to inadequate insulin production. Diabetic nephropathy (DN), is a critical complication associated with DM, often leading to end-stage renal failure and increased mortality. OBJECTIVE: This meta-analysis aimed to evaluate the association between the ELMO1 rs741301 polymorphism and susceptibility to DN among individuals with diabetes. METHOD: A systematic literature search was conducted for studies published between 2014 and 2024 using Embase, Google Scholar, and PubMed. Eligible case-control studies investigating the association between ELMO1 rs741301 and DN among individuals with DM were selected according to predefined inclusion criteria. Nine case-control studies comprising 880 individuals with DM and 1008 individuals with DN were included in the meta-analysis. RESULTS: The pooled analysis demonstrated a significant protective association between the ELMO1 rs741301 polymorphism and DN under the allelic model (OR = 0.77, 95% CI: 0.67-0.88), recessive model (OR = 0.74, 95% CI: 0.61-0.90), and dominant model (OR = 0.68, 95% CI: 0.53-0.88). In contrast, no statistically significant association was observed under the over-dominant model. CONCLUSION: The findings suggest that the ELMO1 rs741301 polymorphism may be associated with a reduced susceptibility to DN among individuals with DM. These findings provide evidence for a potential genetic contribution of ELMO1 to DN susceptibility and may help improve understanding of the genetic factors underlying diabetic complications. Further well-designed studies in diverse populations are warranted to validate this association.

Humans

Association of ERBB4 and SHBG gene polymorphisms with polycystic ovarian syndrome in South Indian women: a case-control genetic analysis.

INTRODUCTION: Polycystic ovary syndrome (PCOS) is a multifactorial endocrinological disorder with a substantial genetic component. However, the role of genes involved in follicular development and androgen regulation remains incompletely understood, particularly in South Indian populations. This study aimed to evaluate how variations in the ERBB4 and SHBG genes affect PCOS risk. METHODOLOGY: A hospital-based case-control study was conducted among 400 South Indian women, comprising 200 women with PCOS and 200 age-matched healthy controls. Genomic DNA was extracted to study SNPs at ERBB4 (rs2178575 and rs1351592) and SHBG (rs1799941 and rs727428) using ARMS-PCR genotyping. The study compared genotype and allele frequencies between cases and controls while assessing their associations with allelic, homozygous, heterozygous, dominant, recessive, and over-dominant genetic models. Genotyping accuracy was confirmed by re-genotyping and Sanger sequencing of a subset of samples. RESULTS: The ERBB4 rs2178575 polymorphism demonstrated a significant association with PCOS, as the AA genotype and A allele combination increased risk across all three genetic models, including homozygous, recessive, and allelic models. The ERBB4 rs1351592 variant was associated with 3-fold higher risk of PCOS in heterozygous and GC carriers. The SHBG rs1799941 polymorphism showed a significant link to PCOS through its effects on heterozygous and allelic states, whereas rs727428 displayed no significant connection due to its monomorphic distribution. CONCLUSION: These findings suggest that polymorphisms in ERBB4 and SHBG may contribute to PCOS susceptibility in South Indian women in a locus- and model-specific manner, revealing the intricate genetic structure that defines this medical condition.

Humans

Performance comparison of rapid and native barcoding methods for Oxford Nanopore sequencing of Poliovirus Viral Protein 1 (VP1) amplicons.

Accurate and timely sequencing of poliovirus is critical for global eradication efforts, particularly for molecular epidemiology based on the typing region of the genome, viral protein 1 (VP1). While Oxford Nanopore Technologies (ONT) sequencing has expanded capabilities for poliovirus surveillance, the relative performance of different ONT library preparation methods, including ligation-based (Native Barcoding) and transposase-based (Rapid Barcoding) approaches, has not been systematically evaluated. In this study, we compared rapid barcoding and native barcoding workflows for sequencing VP1 amplicons from 17 type 2 poliovirus-positive samples, each processed in triplicate. Native barcoding generated significantly more sequencing output, producing approximately 2.3-fold greater total read yield than rapid barcoding, and demonstrated higher run-to-run reproducibility (R2 = 0.979-0.998 vs. 0.847-0.929, respectively; p&#x202f;<&#x202f;0.001). In addition, native barcoding generated 80% of the total yield achieved by rapid barcoding within approximately 7&#x202f;h, whereas rapid barcoding required approximately 40&#x202f;h to reach the same output. Despite these differences, both methods produced identical VP1 consensus sequences across all samples, with comparable read quality (median per-base Q-scores of approximately Q17-Q18). Rapid barcoding provided substantial practical advantages, reducing hands-on library preparation time (55 vs. 200&#x202f;min) and per-sample cost ($12.82 vs. $16.54), while simplifying workflow and reducing technical complexity. These findings indicate that sequencing yield may not be a determinant of downstream analytical outcomes for poliovirus VP1 ONT sequencing. Rapid barcoding therefore represents a cost-effective and efficient approach for routine poliovirus surveillance, whereas native barcoding remains advantageous in applications requiring rapid data generation or maximal sequencing depth.

Poliovirus

Introgression shapes the genomic conflict landscape of Malus, providing evidence for a reticulate backbone in a woody crop lineage.

Phylogenomic discordance is widespread across plants, but its evolutionary significance is often obscured when conflict is treated primarily as analytical noise rather than as evidence of underlying processes. In woody lineages in particular, incomplete lineage sorting, introgression, and genome duplication can interact over long timescales to produce complex genomic histories that are not adequately summarized by a strictly bifurcating tree. Here, we use Malus as a model woody genus to investigate how these processes structure conflict across a genus-scale, accession-based phylogenomic framework. Using broad taxon sampling, hundreds of nuclear loci, plastid genomes, and genome-wide SNP summaries, we reconstruct a robust nuclear backbone for sampled Malus lineages and evaluate where discordance is concentrated and which processes best explain it. Nuclear analyses resolve eight major clades, whereas conflict is non-random and localized to recurrent hotspots rather than evenly distributed across the tree. Cytonuclear discordance is similarly concentrated, especially around Clade H, represented by sampled accessions of M. tschonoskii, where localized plastid-nuclear disagreement is consistent with candidate plastid capture or organellar introgression. Multiple complementary analyses further indicate that the strongest conflict is not explained by ILS alone, but instead reflects lineage-structured introgression, while polyploid complexes represent additional localized sources of evolutionary complexity. Together, these results provide evidence for a reticulate genomic backbone in Malus and show how integrating nuclear, plastid, and genome-wide conflict analyses can help distinguish background discordance from process-specific signals in woody plant radiations. Several lineage-level reticulation hypotheses identified here should now be tested with broader population-level sampling and curated reference accessions.

Malus

Comprehensive analysis of mRNA-microRNA-lncRNA expression profiles in post-traumatic elbow heterotopic ossification using RNA sequencing and experimental validation.

BACKGROUND: This study aimed to profile the molecular signatures of post-traumatic elbow heterotopic ossification (HO) to identify key regulators and potential therapeutic targets. METHODS: Total RNA from post-traumatic elbow HO tissues (n=4) and normal bone tissues (n=6) was subjected to high-throughput sequencing to identify differentially expressed mRNAs (DEGs), microRNAs (DEMs), and lncRNAs (DELs). Bioinformatics analyses included Gene Ontology (GO), Kyoto Encyclopedia of Genes and Genomes (KEGG) pathway enrichment, protein-protein interaction network construction, and transcription factor (TF)-microRNA-mRNA network analysis. The expression trends of four most upregulated and four most downregulated DEGs were validated by real-time quantitative reverse transcription polymerase chain reaction (qRT-PCR). RESULTS: We identified 2,138 DEGs, 40 DEMs, and 905 DELs. DEGs were significantly enriched in biological process "bone mineralization," cellular component "plasma membrane," molecular function "integrin binding," and pathways including PI3K-Akt, NF-&#x3ba;B, JAK-STAT, and TNF signaling pathways. Hub genes with high connectivity included MMP9, IL6, MMP3, CTSK, and BGLAP. Integrated network analysis highlighted the transcription factor JUN and key microRNAs (hsa-miR-124-3p, hsa-miR-548c-3p, and hsa-miR-135b). The qRT-PCR results confirmed the expression trends of selected DEGs. CONCLUSIONS: This study, for the first time, profiled the differentially expressed mRNAs, microRNAs, and lncRNAs in post-traumatic elbow HO using high-throughput RNA sequencing. These findings provide valuable insights into the molecular mechanisms of HO following elbow trauma. The identified hub genes (MMP9, IL6, MMP3, CTSK, and BGLAP), key TF (JUN), and key microRNAs (hsa-miR-124-3p, hsa-miR-548c-3p, and hsa-miR-135b) may serve as potential therapeutic targets for preventing and treating post-traumatic elbow HO.

Humans

Diagnostic value of plasma cell-free DNA metagenomic next-generation sequencing in patients with suspected infections and exploration of clinical scenarios-a retrospective study from a single center.

BACKGROUND: Plasma cell-free DNA metagenomic next-generation sequencing (mNGS) is a non-invasive comprehensive method for the etiological diagnosis of various infectious diseases. However, research on the early diagnosis and real-world clinical impact of plasma mNGS in patients with suspected infection are still limited. MATERIALS AND METHODS: This study retrospectively included 140 patients with suspected infections who underwent early plasma mNGS and conventional culture testing. Referring to the clinical diagnosis of infectious diseases, the diagnostic performance of plasma mNGS and culture tests was compared, and the application scenarios and clinical effects of plasma mNGS were evaluated. RESULTS: The positive rate of plasma mNGS was significantly higher than that of culture methods (55.71% vs 25.10%, p&#x2009;<&#x2009;0.001) and blood cultures (55.71% vs 12.86%, p&#x2009;<&#x2009;0.001). Regarding clinical diagnosis, the sensitivity of plasma mNGS was significantly higher than that of culture (58.27% vs 37.80%, p&#x2009;=&#x2009;0.002). The combination of mNGS and culture achieved a higher detection sensitivity (69.29%), especially in patients with multi-site co-infections (73.68%) and blood infections (73.17%). Plasma mNGS demonstrated higher sensitivity in patients with procalcitonin (PCT) index > 5&#x2009;ng/ml or human neutrophil lipocalin (HNL) index > 200&#x2009;ng/ml. In terms of treatment, a total of 69 patients (54.33%) benefited from plasma mNGS. CONCLUSION: This study highlights the significant improvement in pathogen detection performance by combining conventional culture with plasma mNGS detection, especially in patients with multi-site co-infections and blood infections. Early use of plasma mNGS as an adjunct to culture can better guide clinicians to initiate appropriate anti-infective therapy.

Humans

Whole-exome characterization of host genetic variation in HIV-associated genes across the high-prevalence Mizo population, Northeast India.

BACKGROUND: The Mizoram state of Northeast India has one of the highest HIV prevalence rates in Asia, yet the host genetic factors influencing HIV susceptibility in this Tibeto-Burman population remain uncharacterised. METHODS: We performed whole-exome sequencing using Illumina NovaSeq 6000, mean coverage 100X on 76 HIV-negative Mizo individuals. Variants were called using GATK HaplotypeCaller v4.3 against GRCh38p14, annotated with ANNOVAR, and filtered using hard-quality thresholds (QD&#xa0;&#x2265;&#xa0;2, SOR&#xa0;&#x2264;&#xa0;3, MQ&#xa0;&#x2265;&#xa0;40, DP&#xa0;&#x2265;&#xa0;10, GQ&#xa0;&#x2265;&#xa0;20). The allele frequencies were compared against gnomAD v2.1.1 population databases. Hardy-Weinberg equilibrium was assessed using the Wigginton exact test with Bonferroni correction. RESULTS: Post-quality filtering resulted in 12,011 sample-variants across 2,821 unique positions from 36 HIV-associated loci (33 protein-coding genes, 2 chemokine ligands, and 3 lncRNA targets). Of these, 784 observations (51 unique positions) were high-impact nonsynonymous or loss-of-function variants. ADAR rs2229857 (p.K384R, NM_015840) was the most frequently observed variant (Mizo carrier frequency&#xa0;=&#xa0;0.895; 95% CI: 0.806-0.946). CXCR1 rs16858808 (p.R335C) showed the greatest population enrichment (Mizo carrier frequency&#xa0;=&#xa0;0.197; 95% CI: 0.123-0.300; 7.65-fold carrier-frequency enrichment versus gnomAD South Asian; CADD&#xa0;=&#xa0;15.60). Sixteen of 20 tested variants deviated from Hardy-Weinberg equilibrium after Bonferroni correction (p&#xa0;<&#xa0;0.0025), predominantly showing excess homozygosity consistent with the endogamous Mizo population. The protective variant CCR5-&#x394;32 was absent in all the 76 individuals tested. CONCLUSION: This first whole-exome characterization of HIV host genes in the Mizo population identifies CXCR1 rs16858808 as the most population-enriched functional variant and reveals a pervasive endogamy signature. These findings provide a population-specific genetic framework for future HIV susceptibility studies and ART pharmacogenomics research.

Humans

Comparison of paralog identification methods and their impact on species tree topologies in target capture phylogenomics within the Sindora clade (Detarioideae: Leguminosae).

Target capture is a common method of generating high throughput DNA sequencing data for phylogenetic reconstruction of species relationships, for which single copy genes are usually most informative. However, a pervasive problem with target capture is that putatively single copy genes may in fact be paralogs resulting from gene duplication, which are problematic for phylogenetic inference because their evolutionary history may differ from the divergence history of species. Here, we use as a case study a target enrichment dataset of 88 species of Detarioideae (Leguminosae) with a focus on the Sindora clade to examine approaches for handling paralogs, including the built-in paralog handling functions in HybPiper and CAPTUS, plus subsequent steps using Putative Paralog Detection and the tree-based Yang & Smith orthology inference approach. We compare the paralogs flagged using these methods and verify their performance with BLAST mapping against a reference genome sequence of Sindora glabra, and then subsequently compare the species tree topologies produced across these methods. Our comparisons of paralogs flagged across the Sindora clade show that the Putative Paralog Detection pipeline was the most accurate in identifying paralogs in terms of its similarity to the BLAST mapping, followed by the built-in paralog identification function of CAPTUS. However, the results we recovered for the Detarioideae subfamily suggest that the largest differences in species tree topology resulted from the use of paralog-filtered alignments (such as with the Putative Paralog Detection pipeline and the Yang & Smith orthology inference approaches) rather than just by removing the sequences of identified paralogous genes. This was the true for HybPiper-assembled datasets but was not seen in CAPTUS-assembled datasets. In all comparisons, the topological differences caused by different paralog handling methods tended to be confined to clades where processes such as hybridisation and introgression are prevalent. Our study provides a roadmap to establish the best approach to identify, eliminate or separate paralogs in the absence of a chromosomally contiguous reference genome for a study group, and highlights the importance of careful data inspection and processing in addition to understanding the extent of paralogy and paralog characteristics (e.g. sequence divergence between copies) for their study group.

Phylogeny

Integrative genomic and transcriptomic analyses identify key regulators of skin pigmentation in Larimichthys crocea.

The yellow body coloration of large yellow croaker (Larimichthys crocea) constitutes a crucial economic trait, yet its underlying genetic regulatory mechanisms remain poorly understood. This study systematically elucidated the molecular basis of body color variation by integrating genome resequencing and skin transcriptome analyses, combined with the contextual analysis of key pigmentation-related genes and phenotypic histological validation. 200 phenotyped individuals (including yellow-selected lines, F1 progeny, and normal control groups, all derived from a well-characterized aquaculture stock) identified 39 significantly associated SNPs (-log&#x2081;&#x2080;(P)&#xa0;&#x2265;&#xa0;6), mapping to multiple candidate genes. These genes were significantly enriched in pathways related to pigment deposition (GO:0033059), melanosome organization (GO:0032438), melanogenesis, and tyrosine metabolism. Cross-developmental stage transcriptome analysis revealed 2395 differentially expressed genes (DEGs). Multi-omics integration identified eight overlapping candidate genes, including tyrp1, slc45a2, oca2, and dgat2, among which tyrp1 was prioritized for in-depth validation based on its core regulatory role in eumelanin synthesis, significant SNP association signal, and consistent downregulation in transcriptomic data. Experimental validation demonstrated that the g.895C&#xa0;>&#xa0;T mutation in exon 2 of tyrp1b was strongly significantly associated with the yellow phenotype: the frequency of mutant genotypes (TT/CT) reached 92.86%in the yellow-selected group, whereas the control group exclusively exhibited the wild-type genotype (CC). qPCR confirmed significantly downregulated tyrp1b expression in the skin of yellow individuals, consistent with the transcriptome trend. Histological and stereomicroscopic observations of skin tissues further validated the physiological basis of the yellow phenotype, revealing a significant reduction in melanophore number and abnormal melanosome morphology in yellow-phenotype individuals, accompanied by increased xanthophore density. These results suggest that tyrp1b mutation is strongly associated with the yellow phenotype. However, the presence of a wild-type CC individual in the yellow group indicates that this mutation is not strictly required for yellow coloration, suggesting that other genetic or environmental factors may also contribute to the phenotype, Additionally, downregulation of the carotenoid metabolism gene bco2 coupled with upregulation of xdh, together with the functional changes of slc45a2 and oca2, may synergistically promote xanthophore pigment deposition, contributing to the yellow phenotype. As melanin synthesis in large yellow croaker relies on the conserved tyrosinase pathway and transporter proteins, mutations in associated genes (tyrp1b, slc45a2, oca2) represent a primary underlying cause for the loss of melanin-based coloration and transition to a yellow phenotype in L. crocea. These findings provide key molecular targets and a theoretical foundation for molecular breeding of body color in this species, and also enrich the understanding of xanthism regulatory mechanisms in teleosts.

Animals

Quo vadis, BGA? A collaborative EDNAP exercise on the challenges and progress in forensic biogeographical ancestry inference.

There is a broad consensus that forensic tests for the prediction of externally visible characteristics (EVC) and analysis of biogeographic ancestry (BGA) of an individual are technically reliable. However, interpretation of the results and population-specific genotype distribution patterns remains challenging. EVC and BGA analyses provide valuable information for population genetics studies and as investigative leads for criminal cases, as well as for historical and contemporary identification tests. However, inaccurate or incorrect predictions, for example, from subjective bias in the interpretations made, have the potential to misdirect police investigations. The legal situation regarding EVC and BGA testing varies by country: ranging from countries where it is explicitly prohibited, to those without specific regulations on biogeographic ancestry prediction, and others that have already enacted laws governing its use. The reluctance to utilize these analyses is not only due to legal restrictions and data protection concerns, but also to initial limited sets of sufficiently comprehensive forensic DNA assays. Forensic BGA marker panels typically contain up to &#x223c;300 SNPs. This relatively small number of genetic markers, along with limited reference population data, complicates the interpretation of results from donors of unknown origin. This paper presents the results of a collaborative EDNAP study, which, for the first time, evaluated the approach to reporting EVC and BGA data between international laboratories. For the study, DNA from nine individuals with self-reported ancestry was collected and analysed using various forensic panels differing in the number and composition of ancestry-informative markers genotyped, comprising: the Precision ID mtDNA Whole Genome Panel, the VISAGE Basic Tool and the VISAGE Enhanced Tool for Appearance and Ancestry Prediction, and the Ion AmpliSeq&#x2122; PhenoTrivium Panel. To ensure full data protection, all SNP genotypes and uniparental marker haplotypes obtained were not shared with third parties. Instead, the genetic data were analysed using a range of commonly used population analysis software packages. These analysis outcomes were then distributed to twelve European forensic laboratories (both academic and law enforcement institutions), who were asked to prepare reports based on their interpretation of the phenotypes and ancestry they inferred from the analysis data. A questionnaire sent alongside the genetic information, aimed to evaluate which difficulties were encountered by the participants in processing the BGA analysis data they were given.

Humans

O'nyong-nyong virus adaptive mutations in non-structural protein 1 and 3 enhance RNA replication and overcome FHL1 requirement.

Arthritogenic alphaviruses, like o'nyong-nyong virus (ONNV), cause debilitating musculoskeletal diseases and are geographically expanding. To predict their emergence, we seek to better understand evolutionary mechanisms that enable changes in virus tropism. Here, we identify adaptive mutations in the ONNV non-structural proteins (nsPs) that arose during cellular serial passaging and enabled ONNV to infect non-permissive Lunet cells. Using shotgun proteomics, we show that this human hepatoma cell line lacks the four-and-a-half-LIM domain protein 1 (FHL1), an essential host factor in ONNV RNA replication. Individual single nucleotide mutations in the nsP1 ring-aperture membrane-binding and oligomerization domain, the nsP3 macrodomain, and the nsP3 opal stop codon overcome FHL1 deficiency in Lunet cells by enhanced RNA replication. These findings demonstrate how subtle genomic changes in nsPs can profoundly influence alphavirus replication and tropism.

LIM Domain Proteins

Genome-wide characterization of heat shock protein genes reveals thermal stress-responsive candidates in Litopenaeus vannamei.

Heat shock proteins (HSPs) are conserved molecular chaperones involved in protein folding, refolding, aggregation prevention, and degradation of damaged proteins. However, the genomic organization and thermal responsiveness of HSP genes in the Pacific white shrimp (Litopenaeus vannamei) remain incompletely understood. Here, we performed a genome-wide analysis of the HSP gene family and examined its phylogenetic relationships, structural features, duplication patterns, sequence variation, interaction networks, and transcriptional responses to acute heat stress. A total of 34 HSP genes were identified and classified into the HSP90, HSP70, HSP40/DNAJ, HSP60, and small HSP families. Phylogenetic, motif, gene structure, synteny, and subcellular localization analyses revealed evolutionary conservation and structural diversification among family members. Three duplicated gene pairs were identified, comprising two segmental duplications and one tandem duplication. All pairs exhibited Ka/Ks ratios below 1, consistent with purifying selection of varying strength. Sequence analysis identified 295 nonsynonymous single-nucleotide polymorphisms, of which 12 were consistently predicted to be deleterious by multiple algorithms. Protein-protein interaction analysis indicated enrichment of protein-folding and cellular stress-response functions. RT-qPCR analysis showed significant induction of HSPA4, HSP90AA1, TRAP1, BiP, and DNAJA1 after 6, 12, and 24&#xa0;h of exposure to 34&#xa0;&#xb0;C, whereas DNAJC3 was significantly induced only at 12&#xa0;h. All six genes reached their highest transcript abundance at 12&#xa0;h. These findings may provide a genomic framework for HSP genes in L. vannamei and identify candidate genes and variants associated with thermal stress responses.

Animals

Conserved host-exclusive oligonucleotide motifs enriched in pathogenic genes of human oncogenic viruses.

Comparative viral genomics can reveal sequence-level constraints influencing virus-host interactions. Relative minimal absent words (rMAWs) are short oligonucleotide motifs present in viral genomes but completely absent from the host, potentially reflecting selective pressures related to host adaptation and immune evasion. Using the EAGLE algorithm and the GRCh38 human reference genome, we systematically screened for prevalent rMAWs (prMAWs) across six major human oncogenic viruses: Epstein-Barr virus (EBV), hepatitis B virus (HBV), hepatitis C virus (HCV), human papillomavirus (HPV), human T-cell leukemia virus type 1 (HTLV-1), and human herpesvirus 8/Kaposi's sarcoma-associated herpesvirus (HHV-8/KSHV). highly conserved 11- and 12-bp prMAWs were identified in EBV, HBV, HTLV-1, and HHV-8/KSHV, with sequence prevalences ranging from 91.5% to 97.9%. Conversely, no short prMAWs were detected in HCV or HPV, likely reflecting differences in genome architecture, mutation rates, and long-term host adaptation to the human host. Importantly, the identified host-exclusive motifs exhibited non-random genomic distribution and were preferentially embedded within viral genes central to replication, persistence, immune modulation, and oncogenesis, including EBNA-1 (EBV), HBx (HBV), Tax-associated regions (HTLV-1), and lytic replication genes of HHV-8/KSHV. Notably, all detected prMAWs were enriched in GC nucleotides and exhibited marked CpG over-representation, suggesting sequence constraints associated with epigenetic regulation and viral persistence. Collectively, these highly conserved, host-exclusive signatures offer promising, candidates for sequence-directed approaches in the diagnosis, monitoring, and investigation of virus-associated cancers.

Humans

A novel peptide encoded by circTLL1 drives osimertinib resistance in lung cancer by modulating the NT5C2/Ras/PI3K axis.

BACKGROUND: Acquired resistance to osimertinib, a third-generation EGFR tyrosine kinase inhibitor, remains a major clinical challenge in the treatment of non-small cell lung cancer (NSCLC). Although circular RNAs (circRNAs) have been increasingly implicated in drug resistance, most studies have focused on their canonical role as microRNA sponges, while their capacity to encode functional micropeptides remains largely unexplored. This study aimed to identify novel circRNAs involved in osimertinib resistance and to characterize their regulatory functions at the protein level. METHODS: Osimertinib-resistant (OR) NSCLC cell lines were established and validated. High-throughput RNA sequencing was performed to compare the circRNA expression profiles between parental and OR cells. The function of the candidate circRNA was assessed through a series of in vitro and in vivo experiments, including cell viability assays, apoptosis analysis, and xenograft mouse models. Mechanistic investigations involved mass spectrometry, co-immunoprecipitation and western blotting to explore its protein-coding potential and downstream signaling pathways. RESULTS: We identified a novel circRNA, termed circTLL1, that was stably and significantly upregulated in OR-NSCLC cells. Functionally, overexpression of circTLL1 promoted osimertinib resistance, whereas its knockdown restored drug sensitivity both in vitro and in vivo. Mechanistically, we discovered that circTLL1 harbors an open reading frame (ORF) that is translated into a novel 90-amino-acid protein, which we designated circTLL1-90aa. Further investigation revealed that circTLL1-90aa directly interacts with and promotes the degradation of 5'-nucleotidase, cytosolic II (NT5C2), thereby uncoupling nucleotide metabolism from its normal regulatory constraints. The consequent downregulation of NT5C2 leads to elevated GTP levels and leading to the sustained activation of the downstream Ras/PI3K/AKT signaling pathway. CONCLUSION: Our findings unveil a previously unrecognized circRNA/micropeptide/metabolism cascade underlying osimertinib resistance. The identification of the circTLL1-90aa/NT5C2/Ras/PI3K axis not only expands the functional repertoire of the non-coding genome but also provides new insights into the complexity of drug resistance. Given its selective upregulation in resistant cells, circTLL1-90aa holds promise both as a predictive biomarker for treatment stratification and as an actionable therapeutic target, offering a novel strategy to overcome osimertinib resistance in NSCLC patients.

Pyrimidines