Search PubMedSearch

PubMed · 41074031

SMART-RNA-Metavirome: a practical RNA metavirome platform compatible with high-throughput sequencing of both short and long reads.

Abstract

BACKGROUND: The RNA virosphere's extensive diversity and its role in emerging infectious diseases underscore the importance of non-targeted sequencing for identifying unknown or rare pathogens, including co-infections. However, enriching low-abundance viral sequences in RNA metaviromics, particularly in the preparation of cDNA libraries and their compatibility with next-generation sequencing (NGS) and third-generation sequencing (TGS), remains challenging. Therefore, our objective is to develop and systematically assess a practical RNA metavirome methodology specifically tailored for the enrichment of low-abundance viral sequences within samples. METHODS: We developed the SMART-RNA-Metavirome platform, integrating SMART-9n library preparation with NGS and TGS technologies. Total RNA was extracted from two field-collected wild Aedes albopictus pools, along with one laboratory-infected Ae. albopictus pool harboring dengue virus (DENV). This RNA was subjected to reverse transcription using both this optimized protocol and random primer-based methods, followed by high-throughput sequencing on Illumina, Oxford Nanopore, and QitanTech Nanopore technologies. Welch's t-test was employed for comparative analysis of the subsequent RNA metavirome data, specifically to evaluate differences in viral species composition and abundance of viral reads between experimental groups. Furthermore, the effectiveness of this platform was systematically validated via RT-qPCR and SMART-RNA-Metavirome-based Oxford Nanopore sequencing across multiple sample types, including mosquito specimens from DENV-infected Ae. albopictus, serum samples from dengue patients and viral isolates of Japanese encephalitis virus (JEV) and Zika virus (ZIKV). RESULTS: The SMART-RNA-Metavirome platform has been systematically validated to excel in enriching the composition and diversity of the RNA virome (P = 0.04), providing sufficient coverage for the complete reconstruction of viral genomes. When employed in the detection of DENV-infected Ae. albopictus, clinical serum samples, and viral isolates of JEV and ZIKV, this technique exhibits a robust correlation with RT-qPCR (r2 > 0.95). Notably, it demonstrates exceptional sensitivity, ensuring sufficient coverage even in samples of DENV-infected Ae. albopictus with a Ct-value of 35.3, attaining an impressive 99.88% genome coverage. Furthermore, this platform possesses the capability to identify virus species and determine their serotypes. CONCLUSIONS: In our study, the SMART-RNA-Metavirome platform outperforms traditional methods, enriching RNA virome composition and diversity, enabling practical compatibility with both NGS and TGS technologies. It demonstrates significant proficiency in detecting both known and unknown arboviruses, even in low-titer samples such as those from wild mosquitoes and clinical sera. This platform facilitates comprehensive monitoring, risk assessment, and early warning of RNA virus transmissions, enhancing our understanding of RNA virome diversity and ecological patterns.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Xiaohua Liu, Ziyao Li, Xiang Guo, Liu Ge, Minling Hu, Qing He, Xiaoqing Zhang, Ziqing Feng, Yuji Wang, Lingzhai Zhao, Shu Zeng, Wenwen Ren, Haiyang Chen, Chunmei Wang, Rangke Wu, Wei Zhao, Fuchun Zhang, Xiao-Guang Chen, Xiaohong Zhou. 2025-10-10. SMART-RNA-Metavirome: a practical RNA metavirome platform compatible with high-throughput sequencing of both short and long reads.. https://doi.org/10.1186/s40249-025-01371-z

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Systematic performance evaluation and application validation of an end-to-end NGS workstation.

Next-generation sequencing (NGS) library preparation is a core component of precision genomics, but it is commonly constrained by inefficiency, variability, and low throughput of manual protocols. To address these limitations, we developed and systematically evaluated a fully automated NGS workstations and further validated its performance across representative application scenarios. The automated system reduced total processing time from 8 to 10 to 4–6 h. At the same time, it maintained similar performance in pre-library metric, including DNA yield and fragment size, as well as post-capture sequencing metrics (Q30 > 90%, mapping rates > 95%, on-target rates 85–90%). The duplication rate was reduced to 5–8%, compared with 10–15% for manual methods, indicating increased library complexity. Bioinformatic evaluation of inter-species read mapping showed minimal cross-contamination, with a maximum contamination ratio of 0.0003%, indicating effective sample isolation in the automated workflow. High concordance in variant detection was observed between automated and manual workflows. Overall, this automated workstation provides a standardized and reproducible workflow that supports scalable precision genomics applications.

High-Throughput Nucleotide Sequencing

RUMINA: high-throughput deduplication of unique molecular identifiers for amplicon and whole-genome sequencing with enhanced error correction.

MOTIVATION: Unique molecular identifiers (UMIs) are widely used in next-generation sequencing to enable accurate molecular counting and error correction. However, challenges remain in accurately collapsing UMI clusters, especially when read counts are low or sparse read clusters arise from barcode sequencing errors. RESULTS: We present RUMINA, a Rust-based pipeline for UMI-aware deduplication and error correction, optimized for both amplicon and shotgun sequencing. RUMINA supports multiple UMI cluster strategies, alongside majority-rule read selection independent of mapping quality, as well as discrete handling of 1-2 read clusters, paired-end merging, and read-length stratification. Benchmarking using simulated HIV population sequencing data and real-world iCLIP and TCR datasets showed that RUMINA improves ultra-low frequency SNV detection (0.01%-1%), reduces false positives, enhances reproducibility, and processes sequencing data up to 10-fold faster than existing tools. By integrating UMI- and sequence-level correction in a high-performance framework, RUMINA offers a fast, scalable, and robust solution for UMI-enabled sequencing workflows. AVAILABILITY AND IMPLEMENTATION: RUMINA is implemented in Rust and distributed as open-source code and precompiled binaries. Source code and installation instructions are available at https://github.com/greninger-lab/rumina. Documentation associated with this manuscript is available at https://github.com/greninger-lab/rumina_paper.

High-Throughput Nucleotide Sequencing

Enzymes in high-throughput RNA sequencing: Applications and challenges.

High-throughput RNA sequencing provides genome-wide information on the dynamics of RNA in each cell and how the dynamics responds to environmental changes. Next-generation sequencing by the Illumina platform currently provides the highest information output as compared to other platforms. A key component of next generation sequencing of each RNA is the successful end-to-end reverse-transcription into a cDNA strand. This can be highly challenging given the propensity of each RNA to adopt ordered structures and to contain post-transcriptional modifications. While many reverse transcriptase (RT) enzymes have been developed over the years to maximize read-through of an RNA, their processivity and efficiency varies, raising the question of how to select the RT for the experiment at hand. Here, we use tRNA as a model for genome-wide sequencing, as tRNA has a stable secondary and tertiary structure and has a high density and wide variety of post-transcriptional modifications, presenting one of the most challenging problems of sequencing RNA. We compare the efficiency of end-to-end cDNA synthesis of tRNA among several recent RT enzymes and provide a general sequencing workflow that is applicable to most of these enzymes.

High-Throughput Nucleotide Sequencing