Search PubMedSearch

SEARCH · Search PubMed

Results for “Metagenome-assembled genome”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

558 records · Page 5Linked to original sources

Addressing lignin composition and content via Arabidopsis arogenate dehydratase knockout and over-expression genotypes.

Following the down-selection of 14 Arabidopsis thaliana arogenate dehydratase (ADT) knockout and over-expression (OE) genotypes, the most highly contrasting quadruple knockout adt3/4/5/6 and ADT OE genotypes were subjected to proteomics, metabolomics, and scanning electron microscopy (SEM) analyses as needed, with results compared to Columbia wild-type (WT). The basal adt3/4/5/6 stem cross-sections, ∼70% lignin content reduced, exhibited buckled vessel cell walls and partially detached xylary fibers, in contrast to WT and ADT4m/5 m OE genotypes that did not. Anatomical defects primarily resulted from guaiacyl lignin level reductions in vessels with concomitant increased stem syringyl:guaiacyl (S/G) ratios. Phenylpropanoid and various upstream shikimate-chorismate pathway enzyme abundances, as well as specific monolignol oxidases (laccases/peroxidases), generally increased in adt3/4/5/6 at different stem and rosette leaf growth/development stages, relative to WT. Opposite effects were largely observed with the ADT5m OE genotype. By contrast, flavonoid and glucosinolate pathway enzyme amounts varied. Such enzyme abundance increases were overall unproductive as adt3/4/5/6 was unable to restore WT, ADT4 OE, ADT5 OE, ADT5m OE, and ADT4m/5 m OE secondary metabolite (lignin, phenylpropanoid, lignan, flavonoid, phenolic acid, and glucosinolate) levels. Conversely, ADT OE genotypes did not significantly increase programmed lignin levels or alter S/G compositions. In sum, proteomics analyses of adt3/4/5/6 and adt5 'perceived' that lignin and low molecular weight secondary metabolite amounts were not at 'programmed' levels as for WT and ADT OE genotypes but observed increases in relevant pathway protein abundances were futile. Notably though, proteomics analyses did not lead to predicting that lignin and associated biochemical pathways would have reduced metabolite levels, relative to WT and ADT OE genotypes. Genotype adt3/4/5/6, possibly the highest lignin level reduced genotype reported, did not utilize other phenolics to compensate. By contrast, the differential temporal and spatial deposition of cell wall oxidases again indicate the exquisite control over lignin deposition, and our lack of knowledge of precise lignin structure and assembly in subcellular regions of the lignified cell walls.

Lignin

Emergence of Babesia naoakii infection in Indonesian domestic cattle, a new host record in water buffaloes, and characterization of complete mitochondrial protein-coding genes.

Babesia (B.) naoakii, previously referred to as Babesia sp. Mymensingh, is a recently characterized tick-borne haemoprotozoan parasite of cattle. In Indonesia, we first reported its presence in 2022 from clinically affected cattle in Central Java. To investigate the wider epidemiology of this neglected ruminant-associated Babesia species, we surveyed apparently healthy cattle (Bos indicus) and water buffaloes (Bubalus bubalis) across three districts of Java, Indonesia. A PCR assay targeting the B. naoakii-specific apical membrane antigen 1 (ama1) gene detected the parasite occurrence in 34.39% of assessed cattle (87/253; 95% CI: 28.80-40.44%) and 30.77% of water buffaloes (12/39; 95% CI: 18.47-46.52%). These results represent the first record of B. naoakii infection in water buffaloes in the country and confirm that the parasite circulates in subclinically infected bovine hosts. To characterise this apicomplexan parasite further at the molecular level, we assembled in full length the three mitochondrial protein-coding genes (PCGs): cytochrome c oxidase subunits 1 (cox1) and 3 (cox3), as well as cytochrome b (cytb). These genes were reconstructed by next-generation sequencing of blood DNA collected during the acute haemolytic-phase of B. naoakii infection, from calves that subsequently succumbed to the disease in the endemic area. Phylogenetic analyses of the concatenated amino-acid sequences of cox1, cox3, and cytb placed the Indonesian isolates within a well-supported monophyletic clade, distinct from all previously characterised ruminant-associated Babesia species and sister to the Babesia bigemina/Babesia ovata lineage. This placement confirmed species identity and reinforced the genetic distinctiveness of B. naoakii in Indonesia. Notably, although B. naoakii circulates in peripheral blood and mirrors the diagnostic behaviour of the mild pathogen B. bigemina, its clinical impact more closely resembles that of the severe pathogenic B. bovis, particularly in young animals. This diagnostic-clinical discordance highlights the need for B. naoakii-specific molecular surveillance and species-level differentiation in regions of co-endemicity. Given the high prevalence in subclinically B. naoakii-infected adults, the documented severity of babesiosis in calves, and the potential for substantial economic losses, broader epidemiological investigations and species-specific control measures for B. naoakii are urgently performed. The same holds true for future epizootiological investigations of underdiagnosed B. naoakii-infections possibly circulating in Indonesian endemic ruminant bovids such as the banteng (Bos javanicus), the lowland anoa (Bubalus depressicornis) and the tamaraw (Bubalus mindorensis).

Animals

Longitudinal whole-genome analysis of bluetongue virus identifies conserved serotype-specific genomes and distinct genomic constellations within a Colorado sheep flock (2021-2023).

Bluetongue virus (BTV) is a segmented double-stranded RNA virus of ruminants transmitted by Culicoides spp. biting midges. Although the genome consists of ten segments, classification into serotypes is primarily based on genome segment 2. However, reassortment among genomic segments is a major driver of BTV evolution and diversity. This study used longitudinal whole-genome sequencing to characterize BTV genomes collected from 2021 to 2023 within a single sheep flock in Colorado, where multiple serotypes co-circulate. Whole-genome sequences were generated from fourteen blood samples representing four serotypes: BTV-6, -11, -13, and -17. Longitudinal sampling identified multiple BTV serotypes within individual sheep across consecutive years. Tanglegram analysis comparing segment phylogenies to the segment 2 tree demonstrated incongruent topologies across all genomic segments, suggestive of reassortment or the circulation of distinct genomic constellations. Nucleotide-level comparisons revealed high sequence homology among same-serotype samples from the same year, while the greatest genetic divergence was observed among BTV-17 genomes collected in different years. Additionally, all BTV-13 genomes contained a previously undescribed nonsynonymous substitution in segment 10 predicted to extend the encoded protein by three amino acids. Together, these findings demonstrate that highly conserved BTV genomes and distinct genomic constellations can be detected at the flock level across multiple years. This longitudinal whole-genome approach reveals the genetic complexity of endemic BTV populations, including novel variants and genomic patterns consistent with reassortment that are lost with conventional serotyped-based approaches, highlighting the need to integrate whole-genome characterization into endemic BTV monitoring programs.

Animals

Genome-wide scans reveal candidate genes associated with wing morph differentiation in Tetrix japonica.

Wing dimorphism is an important dispersal-related trait in insects, but its genomic basis remains poorly understood in pygmy grasshoppers. Here, we integrated genome-wide single-nucleotide polymorphism (SNP) analyses, population structure inference, selection scans, and functional annotation to investigate genomic differentiation between long- and short-winged Tetrix japonica. Principal component analysis (PCA), ADMIXTURE, and phylogenetic analyses revealed weak genome-wide separation between morphs, indicating differentiation on a largely shared genetic background. Genome-wide scans based on the fixation index (FST), nucleotide diversity ratios, and Tajima's D, using 50-kb non-overlapping windows and empirical top-5% outlier thresholds, identified multiple candidate regions across seven chromosomes. The broader long- and short-winged candidate sets spanned 9.35 Mb and 9.37 Mb and directly overlapped 82 and 77 genes, respectively. Candidate genes were associated with signaling/hormone regulation, membrane transport, metabolism, cytoskeletal organization, extracellular matrix structure, and development. Short-winged candidate genes were significantly enriched for ABC-type transporter activity and ATP hydrolysis activity. Because all individuals originated from a single laboratory-maintained population with weak genome-wide structure, these regions should be regarded as candidate loci from a screening-stage analysis that require validation in independent populations and by functional assays, rather than as confirmed targets of selection.

Animals

Analyzing salinity tolerance in grass carp (Ctenopharyngodon idella): Insights from genome-wide association study and genomic selection.

Grass carp (Ctenopharyngodon idella) is one of the most widely cultured freshwater fish species globally. However, the expansion of its farming scale faces severe limitation owing to freshwater scarcity; therefore, the development of strains with greater salinity tolerance is key for expanding production using brackish water resources. To investigate the genetic basis of salinity tolerance in grass carp, a genome-wide association study (GWAS) was conducted using 200 individuals representing extreme phenotypes, namely salinity-tolerant and salinity-sensitive groups. In total, 17 single nucleotide polymorphisms (SNPs) related to salinity tolerance were detected, which were distributed across 11 chromosomes. Through gene annotation, 38 candidate genes were obtained from these loci. Enrichment analysis revealed these candidate genes are primarily implicated in key biological processes, including osmotic regulation, energy metabolism, and stress responses. Analyses of different SNP densities revealed that the 5 K SNP density panel can balance prediction accuracy and computational efficiency. The BayesA model achieved the highest prediction accuracy under the GWAS_Evenly selection strategy, with substantial reductions in mean absolute error and mean square error. This study reveals the genetic mechanisms of salinity tolerance in grass carp, which might be optimized through genomic selection, and provides insights for selectively breeding new varieties with greater salinity tolerance.

Animals

Genomic characterization and pathogenicity of ruminant Listeria monocytogenes isolates in a murine oral infection model.

Listeria monocytogenes is a major foodborne pathogen; its ruminant isolates display zoonotic characteristics, causing similar clinical signs in humans, including abortion and encephalitis. However, data on whole genome sequencing and pathogenicity of ruminant L. monocytogenes isolates remain sparse. This study aimed to analyze the genotypic characteristics of L. monocytogenes isolates from ruminants with listeriosis. Furthermore, we assessed the in vivo pathogenicity of four ruminant L. monocytogenes isolates, characterized via whole-genome sequencing-based genetic clustering, in orogastrically inoculated mice. The isolate LM18 (serotype 1/2b, ST224, SL6178) had the lowest lethal dose compared to the other three isolates including previous hypervirulence type (serotype 4b, ST1, SL1) and caused secondary bacteremia in lungs, with sustained bacterial loads in the spleen and liver. Genomic (listeria pathogenicity island -1 and -3) and virulence gene (actA and llsX) mutation analyses associated with virulence suggested from well-recognized studies could not elucidate the virulence of the isolates. SSI-1, which only exists in the isolate LM18 (serotype 1/2b, ST224, SL6178), may help L. monocytogenes survive in the gastrointestinal environment, thereby affecting its virulence. Further research should investigate the role of SSI-1 in the pathogenicity of L. monocytogenes. Moreover, additional studies utilizing larger datasets of ruminant isolates are required to validate our genotypic characterization and to obtain a comprehensive picture of further genotypic differences crucial for L. monocytogenes pathogenicity.

Animals

Whole-Genome Deep Learning Predicts Chemotherapy Response in Colorectal Cancer.

Chemotherapy response in colorectal cancer (CRC) exhibits significant heterogeneity, with current clinical predictors failing to capture complex genomic determinants of resistance. We developed a hybrid deep learning framework integrating convolutional neural networks (CNNs) and bidirectional long short-term memory (BiLSTM) networks to analyze whole-genome somatic mutations, evolutionary conservation, chromatin accessibility, and 3D genome architecture in 2,546 TCGA patients. An attention mechanism identified predictive genomic regions. The model achieved an AUC of 0.92 (95% CI: 0.89-0.94) in cross-validation and 0.88 (95% CI: 0.85-0.91) in independent validation, outperforming clinical models (&#x394;AUC = +0.18, p < 0.001). Key predictors included non-coding variants in TP53, KRAS, and PIK3CA regulatory regions. Triple-positive patients (mutations in all 3 regions) had significantly worse progression-free survival (HR = 4.7, p < 0.001). Our framework enables accurate chemotherapy response prediction and reveals novel non-coding resistance mechanisms, advancing precision oncology in CRC.

Humans

Complete mitochondrial genomes of eight cyclophyllidean tapeworms: genome pattern and phylogenetic analysis.

Cyclophyllidean tapeworms are widespread parasites of significant medical and veterinary importance. However, mitochondrial (mt) genomic resources for cyclophyllideans from China, particularly those recovered from wildlife hosts, remain comparatively limited. In this study, we sequenced and characterized the complete mt genomes of eight cyclophyllidean isolates collected from diverse wild and domestic hosts in China, including two Hymenolepis sp. isolates and two Raillietina sp. isolates from China, and four additional isolates of previously sequenced Taenia species. The circular mt genomes ranged from 13,387 to 14,021&#xa0;bp in length, encoding 36 typical genes with variable non-coding regions. Comparative analysis revealed highly conserved gene composition and mostly conserved mt architecture, with localized rearrangement patterns detected among the cyclophyllidean lineages examined. In particular, all sampled Taeniidae exhibited a consistent trnL1-trnS2 arrangement, whereas the examined non-Taeniidae families showed the trnS2-trnL1 arrangement, confirming and extending, across additional wildlife-associated isolates, a previously proposed family-associated gene-order marker within Cyclophyllidea. Phylogenetic analyses based on concatenated amino acid sequences of the 12 protein-coding genes placed the eight isolates within their expected families, in topologies broadly consistent with previous mitogenomic studies. These data provide additional Chinese mitogenomic references, especially for underrepresented wildlife-associated isolates, and support family-associated gene-order patterns in Cyclophyllidea.

Animals

Plant species identification by genome skimming across the vascular plant tree of life.

Accurate species identification is essential for biodiversity conservation and sustainable use, yet standard plant DNA barcoding often fails to achieve species-level resolution. We present a large-scale empirical evaluation of genome skimming as a tool to improve plant species discrimination. Using standardised data from 1969 individuals representing 475 species from 32 genera across major lineages of the vascular plant tree of life, we compare conventional plastid + internal transcribed spacer (ITS) barcodes with genome skimming approaches. Standard barcoding using rbcL, matK, trnH-psbA and ITS resolved about half of species (49.3%), with six genera showing <&#x2009;25% species discrimination. By contrast, genome skimming enabled the recovery of complete plastid genomes, yielding 57.6% species discrimination. It also generated sufficient nuclear genomic data for additional resolution from k-mer analysis, achieving 66.8% species discrimination - an average gain of 17.5% over standard barcodes - while eliminating cases of extreme failure (<&#x2009;25% resolution). The recovery of complete plastomes and ribosomal DNAs from genome skims also ensures backward compatibility with existing barcode datasets. Our results demonstrate that genome skimming provides data that substantially improves species-level resolution across diverse plant lineages and offers a scalable, high-throughput approach for building comprehensive reference resources to support global biodiversity initiatives.

DNA Barcoding, Taxonomic

Comparative genomic epidemiology of food- and patient-derived diarrheagenic Escherichia coli from sentinel surveillance in Southeast China.

Diarrheagenic Escherichia coli (DEC) remains an important foodborne pathogen, yet long-term comparative genomic surveillance data jointly characterizing food-derived and patient-derived isolates remain limited. This surveillance-based comparative study integrated antimicrobial susceptibility testing and whole-genome sequencing to characterize diarrheagenic Escherichia coli isolates recovered from food and patient sources in Lishui, Southeast China, during 2018-2025, with emphasis on occurrence, resistance profiles, genomic backgrounds, and plasmid replicon-associated features. Antimicrobial susceptibility testing was performed for 258 selected isolates, and whole-genome sequencing was conducted for a curated analytical subset of 204 isolates. The sequenced subset was used for diversity-oriented comparative genomic analysis rather than for unbiased prevalence estimation of the entire DEC collection. EAEC predominated in both sources, although food-associated occurrence was heterogeneous across categories, with the highest recovery rate observed in raw meat. Patient-derived isolates showed a broader overall resistance burden, whereas food-derived isolates retained substantial resistance to tetracycline, chloramphenicol, and florfenicol. Phylogenetic analysis showed partial overlap in genomic backgrounds between food-derived and patient-derived isolates, while representative resistance determinants displayed both broadly distributed and lineage-enriched patterns. Replicon-based plasmid profiling identified 42 plasmid types, including 12 detected in both sources, with IncF-related replicons predominating among these shared profiles. Several food-derived isolates carried multiple plasmid replicon types that were also observed in patient-derived isolates. Overall, food-derived and patient-derived DEC showed partial overlap in genomic backgrounds, resistance determinants, and replicon-defined plasmid profiles within this surveillance setting, while retaining source-associated heterogeneity. These findings should be interpreted as surveillance-based comparative evidence rather than as evidence of direct source attribution or transmission.

Humans

Genomic science and the nurse educator's role: Promoting integration from curriculum to clinical practice.

BACKGROUND: Registered nurses and nurse educators play a critical role in preparing future clinicians to translate genomic discoveries into practice. However, emerging evidence suggests that both groups may lack sufficient knowledge and confidence in genomics, potentially limiting their ability to teach, mentor, and apply genomics in real-world settings. This gap is especially concerning in Aotearoa New Zealand, where the genomic literacy of nurse educators and clinicians remains underexplored. OBJECTIVE: This study aims to: (1) assess nurse educators' genomic literacy and confidence in teaching genomics; and (2) evaluate registered nurses' knowledge and confidence in applying and teaching genomics in clinical practice. DESIGN: Exploratory descriptive qualitative. SETTING: This study was conducted in the greater Auckland area. PARTICIPANTS: A total of 17 participants were recruited using purposive sampling to ensure a diverse range of perspectives across varying levels of teaching experience, disciplinary backgrounds, and exposure to genomic content. METHODS: Data were collected using semi-structured focus group interviews, a method well-suited for generating in-depth discussion and facilitating interaction among participants with shared professional interests. The collected data were analysed using thematic analysis methods. RESULTS: The findings offer insight into the preparedness of New Zealand's nursing workforce to engage with genomic-informed healthcare and inform strategies for integrating genomics into nursing curricula and continuing professional development. Given the interdisciplinary nature of genomic healthcare, these insights may also be relevant to other health professionals-including midwives, pharmacists, and allied health practitioners-who increasingly encounter genomic information in clinical practice and require foundational competencies to support patient care. CONCLUSION: Addressing this educational gap is critical to ensuring that nurses-key facilitators of patient care and public health-are equipped to deliver safe, equitable, and evidence-based genomic healthcare.

Humans

Complete genome sequence of the Anaplasma phagocytophilum clinical isolate NCH-1.

Anaplasma phagocytophilum is an obligate intracellular gram-negative bacterium and etiologic agent of human granulocytic anaplasmosis. A. phagocytophilum genomic sequencing has historically been performed via short-read platforms. Our optimized bacterial isolation protocol combined with Nanopore sequencing produced a single, closed 1,481,805 bp circular A. phagocytophilum strain NCH-1 chromosome.

Anaplasma phagocytophilum

Integrative modeling of the genome structure and dynamics in fission yeast.

Genome organization in the nucleus is highly structured and dynamic. Recent advances in genomic technology have enabled the measurement of genome-wide architecture and locus-specific motion, yielding contact maps and live-cell trajectories. However, these outcomes are derived from different modalities and are not directly comparable, with their quantitative integration being a key challenge. Here we establish a genome-wide live-cell imaging platform in fission yeast Schizosaccharomyces pombe, tracking 131 chromosomal loci, along with the spindle pole body (SPB) and nucleolus, to construct a quantitative map of locus dynamics. By integrating these dynamics with contact data through polymer modeling of Hi-C data, we build a physics-based "digital twin" of the S. pombe genome consistent with the spatiotemporal dynamics of interphase chromatin. We validate it against genome-wide mobility patterns and known architectural features, including centromere and telomere clustering. The model also identifies distinct dynamical regimes: centromere- and telomere-proximal loci relax within [Formula: see text]150 s, whereas the remaining loci relax within [Formula: see text]70 s. We measure semiperiodic dynamics of SPB motion, including a characteristic peak near 225 s and [Formula: see text] fluctuations. We use the model with SPB-directed forcing to show how these low-frequency components propagate through the genome to drive genome-wide chromatin displacements. Together, this predictive physics-based modeling framework integrates genome structure and dynamics to reveal how nuclear mechanical driving forces shape chromosome motion, linking mechanically driven chromatin responses to genome maintenance and regulation.

Schizosaccharomyces

Evaluation of one-step amplicon-based targeted enrichment for SARS-CoV-2 whole-genome sequencing using the Midnight amplicon scheme.

Genomic surveillance proved invaluable during the COVID-19 pandemic for tracking SARS-CoV-2 variants and guiding outbreak responses, underscoring the ongoing need to reduce whole-genome sequencing (WGS) costs and improve workflow efficiency to ensure accessibility in resource limited settings. Here, we evaluated a one-step reverse transcription polymerase chain reaction (RT-PCR) approach using the Midnight V2 primer scheme for targeted amplification of the SARS-CoV-2 genome, assessed its compatibility with Illumina sequencing, and compared its performance to a well-established two-step method. Initially, we determined optimal RT-PCR reaction conditions using the Midnight V2 primer panel for the one-step RT-PCR kit and scaled reaction volumes for both RT-PCR and library preparation. Clinical specimens (n&#x202f;=&#x202f;53) that had undergone routine WGS for surveillance purposes using the established two-step RT-PCR method were compared using the one-step RT-PCR assay. For samples with genome completeness greater than 70%, both methods gave comparable results with similar sequence coverage and 100% concordance for lineage assignment. Further investigation revealed a higher percentage of reads aligning to the SARS-CoV-2 genome with a greater depth of coverage using the one-step method compared to the two-step method. Finally, analysis of scaled one-step and library reaction volumes revealed significant cost savings for samples undergoing WGS. Overall, the results presented here verify the accuracy and reproducibility of one-step targeted amplification and offer an efficient and cost-effective workflow for routine SARS-CoV-2 genomic surveillance.

Humans