A global digital navigator of human health for precision medicine.
Explore the source record for details and available documents.
Biomedical subjects
Publications and source records attributed to Wei Wu.
Explore the source record for details and available documents.
PURPOSE: This study sought to elucidate the possible biological association between BA and lung adenocarcinoma through an analysis of cases in which both lesions coexist within the same specimen. METHODS: In our cohort, the BA and lung cancer components of nine concurrent-type BAs were microdissected using the Millisect system and subjected to whole-exome sequencing (WES). Their histopathological, immunohistochemical, and genomic profiles were comparatively evaluated. RESULTS: Histopathologically, the BA regions of concurrent-type BAs exhibited a classic bilayered architecture, composed of continuous luminal and basal cell layers. The adjacent monolayered concurrent components were diagnosed as adenocarcinoma in situ (AIS, N = 4), minimally invasive adenocarcinoma (MIA, N = 2), and invasive adenocarcinoma (ADC, N = 3). Immunohistochemically, both luminal and basal cells in BA regions expressed thyroid transcription factor 1 (TTF1), albeit with more heterogeneous staining intensity compared to that observed in tumor components. Molecularly, EGFR mutations were the most frequently identified in either BA or tumor components, or in both (Case 9). In BA components, mutations included exon 19 p.S752F, exon 19 deletions (p.L747_T751delinsP and p.E746_T751delinsVP), and compound G719C/S768I mutations. Tumor components harbored exon 28 S1130C, exon 19 indel (p.E746_S752delins), and exon 18 p.G719C mutations. Notably, only three cases demonstrated limited overlap of mutations and copy number variations (CNVs) between the two components. Phylogenetic analysis revealed that six cases shared truncal alterations in genes including KMT2A, PIK3CA, SETD2, MITF, PBRM1, and SRSF3, one case harbored a shared canonical EGFR mutation (p.G719C), with an additional p.S768I alteration uniquely detected in the BA component. CONCLUSION: There is insufficient evidence to support BA as a premalignant lesion for lung adenocarcinoma base on morphological and molecular variables, and they may represent distinct pathological entities.
Rho GTPase-activating protein 10 (ARHGAP10) is recognized as a tumor suppressor, yet the functional impact of its alternative splicing isoforms on breast cancer metastasis remains unclear. This study aimed to elucidate the role and regulatory mechanism of ARHGAP10 exon 21 skipping in breast cancer progression. Our research results indicate that in metastatic breast cancer cells, the full-length isoform ARHGAP10-L is downregulated, whereas the truncated ARHGAP10-S is upregulated. The RNA-binding protein HNRNPA0 directly binds to intron 21 of ARHGAP10 pre-mRNA, promoting exon-21 skipping and ARHGAP10-S production. Functionally, ARHGAP10-L and ARHGAP10-S exert opposing effects on breast cancer cell malignancy: ARHGAP10-L suppresses migration, invasion, and lung metastasis, whereas ARHGAP10-S promotes these aggressive phenotypes. Moreover, ARHGAP10-S exhibits enhanced binding to CDC42 and is associated with increased AKT phosphorylation. In a nude mouse model, HNRNPA0 drove lung metastasis by upregulating ARHGAP10-S. These findings establish the HNRNPA0-ARHGAP10 splicing axis as a key regulator of breast cancer metastasis, in which ARHGAP10-S promotes progression via the AKT pathway whereas ARHGAP10-L acts as a tumor suppressor, highlighting the therapeutic potential of targeting this splicing event to combat metastasis.
Pericentric heterochromatin serves as a fundamental component of eukaryotic chromosomes, endowing specialized genomic architecture with broad functional consequences. Although it is universally marked by H3K9me3 modification, the underlying pericentric DNA sequences diverge substantially across species. Here, by leveraging a transposition reporter system combined with a genome-wide RNA interference (RNAi) screen, we identified a specialized mechanism for recruiting SUV39H methyltransferase to initiate pericentric heterochromatin formation. This pathway depends on a highly ordered complex comprising the Puf68, pre-transfer RNAs (tRNAs), and the primer binding site (PBS). Puf68 binds with high affinity to poly-U tracts in pre-tRNA 3' trailer, forming a Puf68/pre-tRNA complex that subsequently base-pairs with the PBS of nascent long terminal repeat (LTR)-retrotransposons. Through direct interaction, Puf68 recruits Su(var)3-9 to these regions, catalyzing H3K9 trimethylation. Notably, Puf68 is sufficient to initiate de novo heterochromatin assembly both at pericentric and ectopically integrated LTR-retrotransposon regions. Our findings not only uncover a previously unrecognized mechanism of heterochromatin initiation but also resolve a long-standing question of how hosts harness nascent LTR-retrotransposon transcripts.
Cancer management remains fragmented across its continuum, from late-stage diagnosis and salvage therapies to non-personalized surveillance. Here, we present Oncoformer, a unified multimodal transformer model trained on the China Oncology Multimodal Prediction and Surveillance Study (COMPASS) cohort (3.67 million individuals, 17.7 million clinical visits) and validated on independent external cohorts, including the UK Biobank. Oncoformer integrates longitudinal electronic health records with chest X-ray imaging to address multiple clinical tasks: pan-cancer diagnosis (area under the receiver operating characteristic curve [AUROC] = 0.956), future cancer prediction up to 1 year before diagnosis (AUROC = 0.869), tumor stage inference (mean AUROC > 0.90), patient-specific treatment-response forecasting, and recurrence-free survival stratification across ten cancer types (all p < 0.01). Staging predictions were independently validated against postoperative pathological endpoints and shown to converge on core cancer genomic pathways. By translating routine clinical data into a dynamic view of cancer evolution, Oncoformer provides a framework for risk-informed cancer prediction and treatment stratification using routine clinical data.
Duplication of 6p is a rare genetic syndrome of which about 25% have one or more congenital cardiac defects including cardiac septal defects, pulmonary artery hypoplasia and patent ductus arteriosus. We present the first case of a fetus with functional single ventricle and persistent truncus arteriosus during prenatal diagnosis whose genomic analysis revealed a novel pure duplication of 6p25.2-p22.3. The duplication fragment was confirmed to be associated with intrachromosomal insertion from mother via chromosome karyotype. The presentation of this case aims to expand the existing knowledge regarding this rare condition and facilitate its diagnosis in the future. Based on the comparison of cases with 6p duplication syndrome, we notice that the 6p terminal region, especially at 6p25.1 to 6p25.2, could be the critical region associated with heart complications or anomalies of the pulmonary arteries. The duplication of the potential modifier genes, RIPK1 and FARS2, were proposed to be attributed to heart defects in our study and should be further researched.
A major scientific drive is to characterize the protein-coding genome, which is a primary basis for studying human health. But the fundamental question remains of what has been missed in previous analyses. Over the past decade, the translation of non-canonical open reading frames (ncORFs) has been observed across human cell types and disease states1-3, with major implications for biomedical science. However, a key gap in knowledge has been which ncORFs produce small microproteins or alternative protein molecules that contribute to the human proteome. Here we report the collaborative efforts of the TransCODE Consortium4 to produce a consensus landscape of protein-level evidence for ncORFs. We show that about 25% of a set of 7,264 ncORFs gives rise to detectable peptides in a large-scale analysis of 95,520 proteomics experiments. We develop an annotation framework for ncORF-encoded microproteins as human proteins and codify the new conceptual model of 'peptideins' as microproteins that have indeterminate potential as functional proteins. To probe the biological implications of peptideins, we create an evolutionary analysis approach, termed ORF relative branch length (ORBL), and determine that evolutionary constraint is common and associates with observation of ncORF-derived peptides. We then characterize a pan-essential cellular phenotype for one peptidein from the OLMALINC long non-coding RNA. Overall, we generate public research tools supported by GENCODE and PeptideAtlas and advance biomedical discovery for understudied components of the human proteome.
Ribosomal DNA (rDNA) encodes the 18S, 5.8S, and 28S rRNA, accounting for ∼70% of cellular transcription. Despite its essential role and links to cancer and aging, quantifying rDNA instability in mammals remains challenging due to its repetitive organization and inherent heterogeneity. Here, we developed a murine rDNA FISH probe and genomic tools tailored for laboratory mouse strains. The results confirmed rDNA cluster locations, revealed substantial inter- and intra-strain as well as intercellular heterogeneity in rDNA organization within inbred mice and unstressed cells, and identified sources of spontaneous and replication-associated DNA double-strand breaks in the rDNA transcription termination region. Using mouse embryonic stem cells, we showed that BRCA1-mediated homologous recombination promotes rDNA instability, the non-homologous end joining factor XRCC1, but not Ku, suppresses intra-cluster deletions, and ATM kinase preserves rDNA cluster stability. Together, these findings establish a platform and tools for studying rDNA instability in animal models relevant to aging and cancer research.
Mycobacterium tuberculosis complex (MTBC) is distributed globally and has posed a severe threat to human health throughout history. In this study, we analyzed whole-genome data from the four major MTBC sub-lineages prevalent in China (L2.2, L4.2, L4.4, and L4.5) to reconstruct their transmission and expansion histories across East Asia and parts of Central Asia. We found that L2.2 has established a highly connected transmission network centered in Southern China, whereas L4.2 is characterized by cross-border transmission between Central Asia and Western China, and L4.4 and L4.5 exhibit repeated transmission events between Southeast Asia and Southern China. By reconstructing their population histories, we demonstrated that these sub-lineages have experienced multi-stage expansions since the 15th century, accompanied by a recent rapid proliferation of evolutionary clades. These findings reveal that the MTBC epidemic in East Asia may follow a pattern of long-term historical adaptation superimposed with recent concentrated outbreaks, providing potential genomic evidence to inform precise regional tuberculosis control strategies in China.
BACKGROUND: Renal cancer presents a significant global health challenge due to its rising incidence and mortality rates. Often undetected in early stages, it complicates diagnosis and treatment. Current therapies face resistance and limited effectiveness, especially in advanced stages. The diverse subtypes of renal cancer highlight the need for new biomarkers and risk assessment tools for targeted treatments. OBJECTIVE: This study aims to assess the prognostic significance of global DNA methylation (GM) levels in renal cancer, identify new biomarkers, and evaluate the therapeutic potential of the DNA methyltransferase inhibitor decitabine. METHODS: Data on RNA sequencing, gene mutations, DNA methylation, and clinical outcomes were collected from TCGA and GEO databases. We calculated global DNA methylation scores (GMS) and categorized patients into high, intermediate, and low GMS groups. Survival analysis and genomic analyses were conducted to explore the relationships between GMS, clinical outcomes, and tumor characteristics. RESULTS: Higher GMS was identified as an independent prognostic factor associated with worse outcomes in renal cancer. Patients with elevated GMS showed increased mutations, copy number variations, and a more aggressive tumor phenotype. Treatment with decitabine was observed to reduce tumor hypermethylation and downregulate cell cycle pathway activity, indicating potential therapeutic benefits. CONCLUSION: Global DNA methylation plays a significant role in renal cancer prognosis. GMS may serve as valuable biomarkers for prognosis and personalized treatment strategies. Decitabine shows potential efficacy for high GMS patients, particularly through its impact on cell cycle regulation, underscoring the importance of personalized approaches in cancer treatment.
A major scientific drive is to characterize the protein-coding genome as it provides the primary basis for the study of human health. But the fundamental question remains: what has been missed in prior genomic analyses? Over the past decade, the translation of non-canonical open reading frames (ncORFs) has been observed across human cell types and disease states, with major implications for proteomics, genomics, and clinical science. However, the impact of ncORFs has been limited by the absence of a large-scale understanding of their contribution to the human proteome. Here, we report the collaborative efforts of stakeholders in proteomics, immunopeptidomics, Ribo-seq ORF discovery, and gene annotation, to produce a consensus landscape of protein-level evidence for ncORFs. We show that at least 25% of a set of 7,264 ncORFs give rise to translated gene products, yielding over 3,000 peptides in a pan-proteome analysis encompassing 3.8 billion mass spectra from 95,520 experiments. With these data, we developed an annotation framework for ncORFs and created public tools for researchers through GENCODE and PeptideAtlas. This work will provide a platform to advance ncORF-derived proteins in biomedical discovery and, beyond humans, diverse animals and plants where ncORFs are similarly observed.
In recent years, detection technologies based on the CRISPR/Cas12a method have been extensively utilized in the fields of nucleic acid, enzyme, and macromolecule detection, thereby reinforcing their significant role in the detection landscape. Enhancing the simplicity of design, efficiency, and automation of the CRISPR/Cas12a detection system is essential for advancing its application in diagnostics. Recently, we developed an automated CRISPR/Cas12a design system named AutoCORDSv2. This system can process published genomic sequences of pathogenic bacteria in a high-throughput manner and automatically generate conserved and highly specific crRNA sequences, along with primer sequences for target amplification. This capability facilitates the specific and precise design of the CRISPR/Cas12a detection system. In this study, crRNAs targeting the Hantaan virus (HTNV) and Seoul virus (SEOV), as well as RT-PCR primers and RT-RPA primers, were designed using AutoCORDSv2. The experimental results demonstrated that the CRISPR/Cas12a system, automatically designed by AutoCORDSv2, was specific for the detection of both the HTNV and SEOV, with no cross-reactivity observed with other pathogens. The detection sensitivity reached 6 copies/μL (equivalent to 111 copies per amplification reaction), whether measured by a microplate reader or directly observed with the naked eye. The detection results for 50 samples were consistent with those obtained from commercial RT-qPCR kits, indicating high precision. Furthermore, the CRISPR/Cas12a system designed by AutoCORDSv2 can also be utilized for the development of a single-tube detection system with a sensitivity of 42 copies per reaction. This system combined with a 5-min extraction step and RT-RPA, further underscoring its potential for application.