Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Long-read sequencing”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5Linked to original sources

VisPan: real-time visualisation of multiplex amplicon-based sequencing panels for rapid syndromic surveillance and pathogen detection.

MOTIVATION: Infectious diseases persist as a major global public health challenge. Diverse factors, including climate change, globalization, deforestation, human-animal interactions, lifestyle choices, and various biological factors, can contribute to their emergence and reemergence. Rapid detection and characterization of (re)emerging pathogens are therefore critical for effective outbreak management and for enhancing our understanding of epidemics by monitoring the transmission, spread, evolution, and genomics of pathogens. In this context, next-generation sequencing technologies (NGS), particularly long-read platforms such as Oxford Nanopore Technologies (ONT), have opened new avenues for real-time pathogen monitoring. However, the bioinformatics bottleneck remains a challenge, emphasizing the need for efficient, accessible, and user-friendly analysis tools. RESULTS: Here, we present a tool adapted from the RAMPART software that enables real-time data visualisation of multiplex PCR syndromic panels combined with Oxford Nanopore sequencing. This real-time analysis enables rapid pathogen detection, from raw data acquisition to taxonomic assignment, within minutes. The interface offers dynamic visual tracking of the sequencing run and amplicon coverage, facilitating immediate insights during diagnostic workflows. Validation experiments confirmed the system's reliability, accurately identifying all pathogens present in complex clinical or environmental samples. This tool provides an integrated, user-friendly solution for genomic pathogen surveillance in field or clinical settings.

Software↗

TaxTriage: an open-source metagenomic sequencing data analysis pipeline enabling putative pathogen detection.

MOTIVATION: TaxTriage is a comprehensive pathogen identification workflow designed for both short- and long-read untargeted DNA and RNA sequencing data. Combining read classification, mapping, and de novo assembly approaches, putative pathogens are identified through comparisons to curated pathogens and abundance expectations from healthy cohort data. Flexible installation options are enabled using Nextflow™ (NF), including cloud deployment via NF Tower (Seqera Platform) and local installation on a variety of systems, including standalone installations without external internet access. Final analysis summaries are compiled into an Organism Discovery Report, which lists likely pathogens and supporting data, including a custom confidence score. RESULTS: Evaluation of published in silico, clinical, and outbreak datasets identified performance comparable to alternative cloud-based processing pipelines for expected pathogen and co-infection detection with similar sensitivity and increased specificity. To support both public health and veterinary diagnostics communities, customization options have been incorporated to enable improved performance for host species of interest. AVAILABILITY AND IMPLEMENTATION: Source code for TaxTriage is freely available at https://github.com/jhuapl-bio/taxtriage. TaxTriage v2.1.1 has been archived on Zenodo at https://zenodo.org/records/17081354 to permit reproducible analysis as described in this manuscript.

Software↗

Small Copy Number Neutral Intrachromosomal Translocation of PAX6 and Aniridia.

IMPORTANCE: Approximately 5% to 10% of individuals with classic aniridia do not receive a molecular diagnosis after clinical testing for variants in PAX6 and its downstream regulatory region. OBJECTIVE: To apply optical genome mapping (OGM) and long-read whole-genome sequencing (lrWGS) to diagnose an individual with unexplained classic aniridia. DESIGN, SETTING, AND PARTICIPANTS: High-quality DNA was extracted from the blood of a 16-year-old male patient with classic aniridia and prior negative clinical test results that included sequencing and copy number analysis of PAX6 exons and downstream regulatory region as well as genomic analysis via short-read whole-genome sequencing (srWGS) and analyzed using OGM and lrWGS. All analyses were performed in a research laboratory in Wisconsin from January 2019 to September 2025. INTERVENTIONS: OGM and lrWGS. MAIN OUTCOMES AND MEASURES: Identification of a structural variant disrupting PAX6 expression in an individual with classic aniridia, following negative prior testing including srWGS. RESULTS: OGM identified a 55-kb deletion on 11p13 encompassing all PAX6 exons and exon 12 of ELP4, with insertion of this segment into 11q21. lrWGS delineated the exact breakpoints, confirming that the downstream regulatory region, required for normal PAX6 expression, remained at the 11p13 locus. Consequently, the translocated copy of PAX6 at 11q21 is expected to lack expression due to the loss of its essential regulatory elements. CONCLUSIONS AND RELEVANCE: These findings in an individual with classic aniridia harboring an intrachromosomal rearrangement at the PAX6 locus identified by OGM and lrWGS may represent the smallest reported structural variant to separate the PAX6 coding sequence from its downstream regulatory region. This structural variant may have fallen below the detection threshold of srWGS due to its balanced nature and small size, suggesting OGM and lrWGS would be needed for definitive identification.

Aniridia↗

Serotypic and Genomic Diversity of Vibrio anguillarum in Rainbow Trout Farms in Turkey: Implications for Vibriosis Control and Vaccine Candidate Selection.

Outbreaks of vibriosis caused by Vibrio anguillarum are a persistent constraint on rainbow trout (Oncorhynchus mykiss) aquaculture. However, information on the population structure of field strains in Turkey has been lacking. Here, we report the first systematic serotypic, proteomic, and genomic characterization of 23 V. anguillarum isolates collected over 10&#x2009;years from rainbow trout farms located in six major aquaculture regions of Turkey. Serological analyses based on microagglutination, supported by ELISA characterization of hyperimmune sera, identified a clear predominance of serotype O1, whereas isolate V12 exhibited a non-agglutinating, atypical O-antigen profile. Protein profiling (SDS-PAGE) and immunoblotting showed largely conserved whole-cell protein patterns among the isolates, but distinct immunogenic bands at 14, 18, and 40&#x2009;kDa were detected in isolates V18 and V21. Long-read whole-genome sequencing revealed that most Turkish isolates grouped within the global O1 clade, while V12, V25, and V28 isolates occupied more distant branches. Comparative genomics demonstrated a conserved core virulence gene set (RTX toxins, siderophore and iron-uptake systems, motility and adhesion factors, Type VI secretion system), with strain-dependent variation in accessory loci such as anguibactin and T6SS-I. Experimental infections of rainbow trout demonstrated significant differences in virulence among isolates (p&#x2009;<&#x2009;0.05), with the V18 isolate showing high, the V15 intermediate, and the V12 low-mortality rates. By elucidating the relationship among the serotype, immunogenic protein profiles, virulence gene repertoires, and in&#xa0;vivo pathogenicity, this study provides a comprehensive overview of the antigenic and genomic diversity of Vibrio anguillarum isolates from Turkey. Notably, the identification of V18 and V21 as promising candidate strains for further vaccine evaluation, characterized by high virulence and unique immunogenic features, provides a scientific foundation for the development of serotype-specific vaccination strategies to mitigate vibriosis-associated losses in aquaculture.

Animals↗

Novel bacterial hosts and mobile genetic structure of tet(X) variants in tetracycline-contaminated aquatic environment uncovered by culture and long-read metagenomics.

Clinically important tigecycline (3rd-generation tetracycline) resistance tet(X) variants were inferred to have evolutionarily originated from environmental bacteria, and have been recognized among environment, human and animals. However, genetic basis for environmental proliferation and dissemination of tet(X) variants remains ambiguous. This study profiled tet(X) variants at gene, contig, isolate, and community levels in environmental community subjected to long-term stepwise increasing oxytetracycline (1st-generation tetracycline) or tigecycline pressure using long-term microcosm experiments, quantitative PCR, bacterial isolation, whole-genome sequencing, and Nanopore-based long-read metagenomics. We confirmed that both oxytetracycline and tigecycline enriched the abundance of tetracycline resistance genes especially oxytetracycline-enriched tet(X3). Unexpectedly diverse bacterial hosts and genetic structure of tet(X)-positive mobile elements in the environment microbiome were identified using bacterial isolation and long-read Nanopore metagenomics. Pseudomonas defluvii was first reported to carry tet(X3) in the chromosome, forming IS26-tet(X3)-res-ISCR2 circular intermediate to transfer between different DNA molecules. Database mining revealed similar mobile segments have prevailed among animal-derived Acinetobacter species. Unlike the widely reported ISCR2-mediated transfer of tet(X6), we identified a novel mobile multidrug transposon TnAs3 where tet(X6) and class 1 integron co-transferred as its passenger region. Mobile tet(X2)-ere(D)-aadS-erm(F)-blaOXA-347 segment was annotated in Runella, and co-occurrences of tet(X2) and ere(D), aadS, blaOXA-347 were also found in Flavobacterium, Arsenicibacter, Chryseobacterium and Pedobacter. Overall, tetracycline-contaminated aquatic microbiome harboured diverse mobile tet(X)-positive segments which have not yet been acquired by clinical pathogens, and thus served as the genetic pool of tet(X) variants together with indigenous bacterial hosts, especially the newly reported Pseudomonas defluvii. Reducing pollution of older-generation tetracyclines would be a proactive way to mitigate environmental evolution and possible clinical effects of tet(X) variants.

Metagenomics↗

Megamimivirus double-stranded DNA linear genomes flanked by highly diverse terminal inverted repeats.

UNLABELLED: Giant viruses have fundamentally expanded our understanding of virology by challenging the conventional boundaries of both virion size and genome complexity. However, the scarcity of isolates has left many of their unique biological features unexplored. Here, we report the isolation and characterization of four new giant virus species belonging to the subfamily Megamimivirinae, sampled from distinct environments across China. Among these, Megavirus daqingense is the first giant virus isolated from an oil reservoir; it exhibits virion stability under high salinity, chloroform exposure, and elevated temperatures, suggesting fitness adaptations to subsurface conditions. Using a hybrid sequencing approach that integrates short- and long-read technologies, we assembled complete linear genomes for all four isolates, each flanked by long terminal inverted repeats (TIRs). Comparative genomic and synteny analyses identified 29 distinct TIRs from 46 megamimivirus genomes. Gene content within these TIRs was highly diverse, with no orthologous proteins conserved across all repeats. Furthermore, TIR genes experienced weaker purifying selection than those in non-TIR regions (i.e., the genomic regions excluding the TIRs), consistent with their role as drivers of genome plasticity. Notably, we discovered for the first time that identical tRNA genes are shared between TIRs and non-TIR regions of eukaryotic viruses. Collectively, our work provides insights into the structural and evolutionary complexity of megamimiviruses, revealing TIRs as reservoirs of genetic diversity and hotspots for gene transfer, thereby playing a pivotal role in shaping the dynamic architecture of giant virus genomes. IMPORTANCE: Terminal inverted repeats (TIRs) are critical structural elements at the termini of linear genomes essential for fundamental processes such as recombination, replication, and integration across diverse organisms. However, the inherent limitations of short-read sequencing technologies have left the complete structure, diversity, and evolutionary significance of long TIRs in giant viruses unexplored. In this study, we leverage hybrid sequencing and comparative genomic analyses to unveil the complexity of TIRs across the subfamily Megamimivirinae. We demonstrate that TIRs are dynamic genomic hotspots characterized by remarkable gene diversity and unexpected conservation of specific tRNA genes. These findings establish TIRs as key drivers of genome plasticity, serving as hotspots for horizontal gene transfer and genetic innovation. By resolving the long-hidden terminal structures of megamimivirus genomes, this work provides a foundational framework for understanding how TIRs shape the evolution of giant viruses and, more broadly, advances our understanding of genome architecture in large DNA viruses.

Megavirus↗

Chromosome-level genome assembly of Manglietia pachyphylla.

Manglietia pachyphylla, an endangered evergreen tree within the Magnoliaceae family, is renowned for its exceptional ornamental value in landscape horticulture. Despite its classification as a Category II nationally protected plant species in China, the genetic basis of its adaptive traits and conservation priorities remains poorly understood. To address this, we present the first chromosome-scale genome assembly of M. pachyphylla utilizing an integrated approach combining PacBio HiFi long-read and Hi-C chromosome conformation capture sequencing technologies. The assembled genome spans 2.15&#x2009;Gb (contig N50&#x2009;=&#x2009;43.57&#x2009;Mb), exhibiting a heterozygosity rate of 0.78% and repeat content of 78.64%, predominantly comprising long terminal repeat (LTR) retrotransposons (52.86%). Hi-C scaffolding anchored 99.57% of the assembly to 19 pseudochromosomes, achieving a BUSCO completeness score of 96.4%. Annotation revealed 42,505 putative protein-coding genes, with 84.46% of predicted genes were functionally annotated. Phylogenomic analysis positioned M. pachyphylla and Oyama sieboldii clustered together in a well-supported group. This high-contiguity genome assembly enables future investigations into adaptive evolution, functional genomics, and evidence-based conservation strategies for this endangered species.

Chromosomes, Plant↗

New insights on Plasmodium gene expression from direct RNA sequencing.

Oxford Nanopore Technology (ONT) direct RNA sequencing enables the sequencing of native RNA molecules without cDNA conversion. The long-read approach captures full-length reads spanning entire genes and has transformed the study of gene expression in Plasmodium parasites by enabling analysis of untranslated regions, isoforms, and alternative splicing. In addition, ONT provides unique insights into non-coding RNAs, RNA modifications, and polyadenylated tail dynamics, which are expanding our understanding of post-transcriptional regulation in Plasmodium, including processes beyond translational repression in gametocytes and sporozoites. Here, we discuss the past and future applications of direct RNA sequencing in Plasmodium research and highlight its advantages, limitations, and future prospects.

Oxford Nanopore Technology↗

MHASS: Microbiome HiFi Amplicon Sequencing Simulator.

SUMMARY: Microbiome HiFi Amplicon Sequence Simulator (MHASS) creates realistic synthetic PacBio HiFi amplicon sequencing datasets for microbiome studies, by integrating genome-aware abundance modeling, realistic dual-barcoding strategies, and empirically derived pass-number distributions from actual sequencing runs. MHASS generates datasets tailored for rigorous benchmarking and validation of long-read microbiome analysis workflows, including ASV clustering and taxonomic assignment. AVAILABILITY AND IMPLEMENTATION: Implemented in Python with automated dependency management, the source code for MHASS is freely available at https://github.com/rhowardstone/MHASS along with installation instructions. Our code is also published on Zenodo at https://doi.org/10.5281/zenodo.17486364. The data underlying this article are available on GitHub at https://github.com/rhowardstone/MHASS_evaluation/.

Software↗

Complete genome sequence of the Anaplasma phagocytophilum clinical isolate NCH-1.

Anaplasma phagocytophilum is an obligate intracellular gram-negative bacterium and etiologic agent of human granulocytic anaplasmosis. A. phagocytophilum genomic sequencing has historically been performed via short-read platforms. Our optimized bacterial isolation protocol combined with Nanopore sequencing produced a single, closed 1,481,805 bp circular A. phagocytophilum strain NCH-1 chromosome.

Anaplasma phagocytophilum↗

Dynamic Centromeres Under Epigenetic Constraint.

Centromeres are essential chromosomal loci specified epigenetically by CENP-A chromatin, yet they undergo rapid sequence turnover, structural remodeling, and occasional repositioning. In this review, we integrate recent advances enabled by long-read genome assemblies and high-resolution chromatin mapping to synthesize current understanding of centromere organization across taxa. We examine how satellite repeats, transposable elements, molecular drive, and meiotic conflict generate extreme centromere diversity. We further explore how DNA methylation and H3K9me3 heterochromatin constrain CENP-A positioning, stabilize centromeric domains, and shape boundary dynamics during centromere drift, duplication, and de novo formation. Together, these perspectives show how centromeres accommodate evolutionary change while preserving the stringent requirements of faithful chromosome segregation.

Journal Article↗

Population-scale disease-associated tandem repeat analysis reveals locus and ancestry-specific insights.

Tandem repeat (TR) expansions, including short TRs (motifs &#x2264;6&#x2009;bp) and variable number TRs (motifs >6&#x2009;bp), underlie many monogenic disorders, with variable length and sequence influencing pathogenicity, penetrance, severity, and onset. Accurate genotype-phenotype correlation and disease prevalence estimation require characterization beyond repeat length. Here we present a population-scale analysis of 66 disease-associated TR loci using long-read assemblies from 2530 diverse haplotypes from 1265 unaffected donors. Integrating repeat length, motif composition, local ancestry, linkage disequilibrium, and phylogenetic analyses, we reveal extensive locus-, population-, and allele-specific variation shaping disease risk. Up to 8.5% of individuals carry expansions above established pathogenic thresholds, many containing interrupting motifs or sequence structures that attenuate pathogenicity. After excluding alleles from loci with uncertain disease association, non-pathogenic interrupted expansions, and carrier states inconsistent with inheritance patterns, ~4% carried expansions predicted to confer disease risk, largely at adult-onset loci with reduced penetrance. Ancestry-resolved analyses uncover population-specific TR architectures contributing to epidemiological disparities in repeat expansion disorders. Phylogenetic analyses identify conserved ancestral alleles and loci with recent instability. We describe variable linkage disequilibrium patterns and recombination signatures around specific disease-associated TR loci. Our findings emphasize integrating sequence, ancestry, and evolutionary context to understand the complex landscape of disease-associated TRs.

Humans↗

Dogme: a nextflow pipeline for reprocessing nanopore RNA and DNA modifications.

MOTIVATION: Oxford Nanopore (ONT) sequencing allows for the direct detection of RNA and DNA modifications from unamplified nucleic acids, which is a significant advantage over other platforms. However, the rapid updates to ONT basecalling models and the evolving landscape of computational tools for modification detection bring about challenges for reproducible and standardized analyses. To address these challenges, we developed Dogme to automate basecalling, alignment, modification detection, and transcript quantification. Dogme automates the reprocessing of ONT POD5 files by integrating basecalling using Dorado, read mapping using minimap2 and subsequent analysis steps such as running modkit. The pipeline supports three major types of sequencing data-direct RNA (dRNA), complementary DNA (cDNA), and genomic DNA (gDNA). Dogme facilitates detection of diverse RNA modifications supported by Dorado such as N6-methyladenosine (m6A), 5-methylcytosine (m5C), inosine, pseudouridine, 2'-O-methylation (Nm) and DNA methylation, while concurrently quantifying full-length transcript isoforms LR-Kallisto for transcript quantification for dRNA and cDNA. RESULTS: We applied Dogme to three separate mouse C2C12 myoblast replicates using direct RNA sequencing on MinION flow cells. We detected 96&#xa0;603 m6A, 43&#xa0;476 m5C, 8829 inosine, 10&#xa0;055 pseudouridine, and 30&#xa0;320 Nm sites in three biological replicates. The pipeline produced reproducible modification profiles and transcript expression levels across replicates, demonstrating its utility for integrative long-read transcriptomic and epigenomic analyses. AVAILABILITY AND IMPLEMENTATION: Dogme is implemented in Nextflow and is freely available under the MIT license at https://github.com/mortazavilab/dogme, with documentation provided for installation and usage.

RNA↗

Dysregulation of U12-Type Splicing in Lupus Neutrophils.

OBJECTIVE: Neutrophil dysfunction is a hallmark of systemic lupus erythematosus (SLE), but its molecular basis remains unclear. This study explores transcriptional and posttranscriptional changes in low-density granulocytes (LDGs), a proinflammatory neutrophil subset expanded in SLE, focusing on NADPH oxidase (Nox) function and minor intron splicing. METHODS: LDGs and normal-density granulocytes (NDGs) were isolated from patients with SLE and healthy controls (HCs). CYBA (p22phox) expression was evaluated at transcript and protein levels. Nox activity was measured using luminol assays. Bulk RNA sequencing (RNA-seq) and rMATS software were used to assess alternative splicing, particularly of U12-type intron-containing genes. RESULTS: CYBA expression was reduced in SLE LDGs (n&#xa0;=&#xa0;11) compared to SLE and HC NDGs (n&#xa0;=&#xa0;6), with levels resembling those in chronic granulomatous disease neutrophils. SLE LDGs exhibited impaired Nox activity (n&#xa0;=&#xa0;7 SLE, n&#xa0;=&#xa0;12 HC). CYBA is a U12 intron-containing gene, and transcriptomic analysis revealed broad down-regulation of this gene class in SLE LDGs, suggesting minor spliceosome dysfunction. rMATS analysis showed increased U12-type intron retention and widespread splicing defects-including exon skipping and mutually exclusive exon use-in genes such as GBP5, MAEA, and STX10. These abnormalities were validated in an independent long-read RNA-seq data set from SLE peripheral blood mononuclear cells. Importantly, splicing disruptions correlated with disease activity and autoantibody profiles. CONCLUSION: Impaired U12-dependent splicing may contribute to neutrophil dysfunction in SLE, potentially via defective oxidative burst and altered immune regulation. These findings highlight the minor spliceosome as a novel player in lupus pathogenesis.

Humans↗

Optimizing a culture-enriched hybrid metagenomics pipeline to assess the AMR footprint of livestock manure in anaerobic digestate.

The role of environmental samples from livestock production systems, including manure and anaerobic digestate, as reservoirs of antimicrobial resistance genes (ARGs) is likely underestimated because conventional metagenomic approaches can overlook low-abundance ARGs and often lack the resolution to associate these genes with their microbial hosts and co-localized mobile genetic elements (MGEs). We evaluated whether culture-enriched metagenomics (CEMG), with and without antibiotic selection, enhances ARG detection in anaerobic digestate and improves the resolution of ARG-MGE-host associations using hybrid short- and long-read metagenomic assembly. CEMG increased ARG recovery; mean ARG abundance rose from 15.4 counts per million (CPM) in metagenomic fresh digestate (FD) to 124 CPM in CEMG without antibiotics and 160 CPM in antibiotic-selective CEMG. In FD, only 9 unique ARGs were detected, whereas CEMG recovered 112, including ARGs of clinical importance, such as glycopeptide resistance, beta-lactamase genes, and the cfr 23S rRNA methyltransferase conferring cross-resistance to multiple antibiotic classes. Antibiotic selection induced targeted, class-specific shifts in ARG profiles, with ARGs associated with tetracycline resistance consistently enriched across treatments. Hybrid metagenomic assembly resolved the genomic context of 784 ARGs, of which 59.3% were co-localized with at least one class of MGEs, predominantly plasmids and integrative conjugative elements/integrative mobilizable elements. Biocide and metal resistance genes frequently co-occurred with ARGs on the same contigs. Together, these findings demonstrate that antibiotic-selective culture enrichment enhances resistome surveillance by improving detection of low-abundance ARGs, while hybrid assembly provides critical genomic context for assessing their mobility and host associations.IMPORTANCELivestock manure and its byproducts, such as anaerobic digestate, are recognized as important environmental reservoirs of antimicrobial resistance genes (ARGs) and resistant bacteria, yet current metagenomic approaches may underestimate this risk by failing to detect low-abundance but clinically relevant ARGs. Here, we show that integrating culture enrichment with hybrid metagenomics improves ARG recovery and reveals ARG co-localization with mobile genetic elements and putative bacterial hosts. This approach captures a cultivable and condition-responsive fraction of the resistome that is not readily accessible through direct metagenomic sequencing alone, providing a more informative framework for environmental AMR surveillance.

anaerobic digestion↗

Single-cell multi-omics dissects transcript isoform and immune repertoire dynamics in human immunosenescence.

Immunosenescence, a major hallmark of systemic aging, refers to the progressive functional decline of the immune system. This decline not only compromises host defense and immunological memory but also fuels chronic inflammation and tissue degeneration (collectively known as inflammaging). While single-cell RNA sequencing (scRNA-seq) has revealed transcriptomic alterations associated with immune aging, analyses restricted to transcript abundance fail to capture deeper regulatory layers, such as transcript isoform diversity and the remodeling of immune receptor repertoires. To address this limitation, we present a human peripheral immune single-cell multi-omics atlas that integrates gene expression, transcript isoform diversity, and immune receptor repertoires. By combining single-cell full-length transcriptome sequencing (scCycloneSEQ), short-read scRNA-seq, and single-cell immune receptor sequencing (scTCR/BCR-seq), we systematically profiled peripheral blood mononuclear cells (PBMCs) from healthy donors aged 30-40 and 60-70 years. Our analyses uncovered extensive age-related remodeling of immune cell composition, functional states, and TCR/BCR diversity. Notably, we found that CD4+ effector memory T cells exhibited widespread differential isoform usage (DIU), 3'UTR length variation, and a marked reshaping of cytotoxic T lymphocyte (CTL) clonotypes-all of which were closely associated with aging-related inflammation and cellular senescence. This multi-omics atlas delineates key molecular features of immunosenescence and provides a high-resolution resource for deciphering the regulatory architecture underlying immune aging.

TCR/BCR↗

Characterisation of Trichuris incognita n sp in C&#xf4;te d'Ivoire: a morphological, genomic, and genome-wide association with drug sensitivity study.

BACKGROUND: Trichuriasis is a neglected tropical disease that affects up to 500 million individuals and can cause considerable morbidity. For decades, trichuriasis was thought to be caused by one species of whipworm, Trichuris trichiura. The aim of this study was to investigate the origin of differences in response rates to the best available anthelmintic treatment for trichuriasis-a combination of albendazole and ivermectin-in C&#xf4;te d'Ivoire by analysing the parasite population. METHODS: In this morphological, genomic, and genome-wide association study (GWAS) with drug sensitivity we used long-read and short-read sequencing approaches and assembled a high-quality reference genome of Trichuris incognita n sp isolated in a primary interventional study conducted in the Lagunes district of C&#xf4;te d'Ivoire. Children aged 6-12 years were screened between July 14, 2022, and July 31, 2022; children positive for T trichiura on duplicate Kato-Katz smears and with infection intensity of 200 eggs per gram or more were eligible and treated first with albendazole (400 mg) and ivermectin (200 &#x3bc;g/kg) then with oxantel pamoate (20 mg/kg). We constructed a species tree of the Trichuris genus using 12&#x2009;434 orthologous groups. We sequenced individual worms, which were used to confirm the phylogenetic placement and investigate patterns of adaptation through comparative genomic analyses. Finally, we conducted a GWAS to compare albendazole-ivermectin sensitive worms to drug non-sensitive worms. FINDINGS: 670 children were screened, of whom 243 were enrolled and from whom 271 worms were isolated after the first treatment and 827 worms after the second treatment. Sufficient DNA was recovered from 747 worms of which 721 were suitable for further bioinformatic analysis; of these, 179 were albendazole-ivermectin sensitive worms and 542 were drug non-sensitive worms. We present and characterise a new, human-infecting Trichuris species named T incognita n sp, which is morphologically indistinguishable from T trichiura, but forms a distinct phylogenetic clade, closer to Trichuris suis than to the canonical human-infective T trichiura. Comparative genomic analysis of genes suspected to confer resistance to either albendazole or ivermectin in helminths revealed a high number of &#x3b2;-tubulin orthologs, present in the whole population of T incognita n sp, compared with the canonical T trichiura species, but these genes were not associated with a resistant phenotype. The GWAS did not provide conclusive evidence of adaptation to drug pressure within the same species. INTERPRETATION: Our results demonstrate that trichuriasis can be caused by multiple whipworm species, and that differences in response rates might result from species responding differently to drug treatment, rather than from the intraspecies establishment of resistance. This discovery, coupled with the high tolerability of T incognita n sp to albendazole-ivermectin, marks a substantial shift in how we understand and approach whipworm infections. FUNDING: European Research Council.

Trichuris↗

Community-driven updates for comprehensive long-read metagenomics and enhanced binning in nf-core/mag v5.

SUMMARY: nf-core/mag is a reproducible Nextflow pipeline for best-practice metagenomic de novo assembly and binning within the nf-core framework. Here we present a major update that adds support for long-read-only assembly and bin refinement, includes five new binning tools, expands taxonomic classification to viruses and eukaryotes, and improves bin quality evaluation with new tools and latest databases. Through sustained community-driven development spanning seven years and four primary curator teams, nf-core/mag remains actively developed as an open source workflow for metagenomic analysis, benefiting from contributions from across the broader metagenomics, nf-core, and Nextflow ecosystem. AVAILABILITY AND IMPLEMENTATION: The source code of nf-core/mag v5 is available on GitHub (https://github.com/nf-core/mag) under the open source MIT license, with v5.5.0 source code archived on Zenodo (https://zenodo.org/records/21735731). Documentation is viewable on the nf-core website (https://nf-co.re/mag).

Metagenomics↗