Search PubMedSearch

SEARCH · Search PubMed

Results for “Comparative Genomics”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,653 records · Page 4Linked to original sources

Draft genome sequence of Enterococcus casseliflavus strain MBBL_MP4 isolated from healthy bovine milk.

We report the draft genome sequence of Enterococcus casseliflavus MBBL_MP4, recovered from healthy bovine milk. The 3.45-Mbp genome assembly comprises 27 contigs and indicates low pathogenic potential, with no acquired antimicrobial resistance or known virulence genes. This genome provides a valuable resource for the genomic characterization of bovine-associated E. casseliflavus.

Enterococcus casseliflavus

Assembly and Characterization of the First Complete Mitochondrial Genome of Tussilago farfara L.: Insights into Biological Functions and Phylogenetic Relationships within the Asteraceae Family.

Tussilago farfara L., a member of the Asteraceae family, is an economically valuable species due to its edible and medicinal properties. To elucidate the structural characteristics, genetic mechanisms, and evolutionary pathways of the organelle genomes of T. farfara, we sequenced, assembled, and annotated its mitochondrial genome for the first time. The complete mitochondrial genome of T. farfara spans 306,024 bp and contains 33 mitochondrial protein-coding genes (PCGs), 3 rRNAs, and 22 tRNAs. Analysis of the nucleotide substitution rate and genetic diversity revealed that most mitochondrial genome genes may have undergone purifying selection, indicating a slow evolutionary rate and a relatively conserved genomic structure. We further identified 13 fragments of chloroplast-derived DNA integrated into the mitochondrial genome, evidencing intracellular gene transfer. Collinearity analysis showed that Arctium lappa shares the most extensive mitochondrial homologous sequences and the highest sequence similarity with T. farfara. Phylogenetic analysis based on the mitochondrial genome helped to clarify the evolutionary and taxonomic position of T. farfara within the Asteraceae family. The mitochondrial genome sequence of T. farfara provides a valuable genomic resource for species identification and for evolutionary studies within the Asteraceae family.

Genome, Mitochondrial

AI-enabled viral genomics: from virus discovery to host prediction and emerging variant forecasting.

The rapid expansion of metagenomic sequencing has generated vast repositories of viral sequence data that far outpace our capacity to interpret them using conventional approaches. Highly divergent sequences, sparse functional annotation, and taxonomically uneven sampling present fundamental challenges for reference-dependent methods, which lose sensitivity precisely for novel and understudied viruses with high public health relevance. Artificial intelligence (AI) provides a new avenue to address these challenges by enabling predictive inference from viral genomes and proteins while reducing dependence on sequence similarity. In this Review, we discuss representative advances in AI for virus discovery, taxonomic classification and functional annotation, prediction of host range and zoonotic potential, and efforts toward forecasting emerging variants. These advances are transforming viral genomics from a largely descriptive discipline into one with increasing predictive capability. We also critically assess the major challenges that constrain current approaches, including the availability of high-quality and representative datasets, rigorous model evaluation, biological interpretability and responsible governance for increasingly capable AI models.

Artificial Intelligence

Genomic science and the nurse educator's role: Promoting integration from curriculum to clinical practice.

BACKGROUND: Registered nurses and nurse educators play a critical role in preparing future clinicians to translate genomic discoveries into practice. However, emerging evidence suggests that both groups may lack sufficient knowledge and confidence in genomics, potentially limiting their ability to teach, mentor, and apply genomics in real-world settings. This gap is especially concerning in Aotearoa New Zealand, where the genomic literacy of nurse educators and clinicians remains underexplored. OBJECTIVE: This study aims to: (1) assess nurse educators' genomic literacy and confidence in teaching genomics; and (2) evaluate registered nurses' knowledge and confidence in applying and teaching genomics in clinical practice. DESIGN: Exploratory descriptive qualitative. SETTING: This study was conducted in the greater Auckland area. PARTICIPANTS: A total of 17 participants were recruited using purposive sampling to ensure a diverse range of perspectives across varying levels of teaching experience, disciplinary backgrounds, and exposure to genomic content. METHODS: Data were collected using semi-structured focus group interviews, a method well-suited for generating in-depth discussion and facilitating interaction among participants with shared professional interests. The collected data were analysed using thematic analysis methods. RESULTS: The findings offer insight into the preparedness of New Zealand's nursing workforce to engage with genomic-informed healthcare and inform strategies for integrating genomics into nursing curricula and continuing professional development. Given the interdisciplinary nature of genomic healthcare, these insights may also be relevant to other health professionals-including midwives, pharmacists, and allied health practitioners-who increasingly encounter genomic information in clinical practice and require foundational competencies to support patient care. CONCLUSION: Addressing this educational gap is critical to ensuring that nurses-key facilitators of patient care and public health-are equipped to deliver safe, equitable, and evidence-based genomic healthcare.

Humans

Complete genome sequence of the Anaplasma phagocytophilum clinical isolate NCH-1.

Anaplasma phagocytophilum is an obligate intracellular gram-negative bacterium and etiologic agent of human granulocytic anaplasmosis. A. phagocytophilum genomic sequencing has historically been performed via short-read platforms. Our optimized bacterial isolation protocol combined with Nanopore sequencing produced a single, closed 1,481,805 bp circular A. phagocytophilum strain NCH-1 chromosome.

Anaplasma phagocytophilum

Identification of aquaporin (AQP) genes in the noble scallop Chlamys nobilis and characterization of their expression under low-temperature stress.

Aquaporins (AQPs) are transmembrane channel proteins essential for water homeostasis and cellular stress responses. In marine bivalves, their roles in cold tolerance remain poorly understood despite frequent winter mortality events in aquaculture. Here, we identified nine AQP genes in the genome of the economically important noble scallop Chlamys nobilis. Phylogenetic analysis revealed strong conservation with other bivalve AQPs, and structural features, including conserved NPA motifs and ar/R selectivity filters, support their canonical water/glycerol transport functions. Tissue-specific expression profiling showed predominant enrichment in osmoregulatory tissues (gills, intestine) and gonads. Under both chronic and acute low-temperature stress from 23 °C to 9 °C, most CnAQP genes exhibited transient upregulation followed by suppression. Notably, CnAQP4 displayed sustained upregulation, implicating it as a key mediator of long-term cold adaptation. Promoter analysis further revealed abundant cis-elements linked to growth and development as well as immune regulation. Our findings provide the first comprehensive characterization of the AQP family in C. nobilis, highlighting its critical role in maintaining cellular integrity during cold stress and offering molecular targets for selective breeding of cold-tolerant scallop strains.

Animals

Mul-PheG2P: decoupled learning and prediction-space fusion enables robust and interpretable multi-phenotype genomic prediction.

Genomic prediction of multiple phenotypes is crucial in modern plant breeding; however, existing methods struggle with negative transfer and lack interpretability, particularly across high-dimensional small-sample data and diverse species. To address this, we propose Mul-PheG2P, a novel paradigm based on decoupled learning and predictive space fusion. It employs a two-stage design: first training phenotype-specific encoders using genetic data, then decoupling phenotype-specific learning from cross-phenotype aggregation via an interpretable prediction layer. Mul-PheG2P outperforms existing methods across diverse crop datasets, including maize (Zea mays), wheat (Triticum aestivum), and tomato (Solanum lycopersicum). It provides a multi-scale interpretability chain: at the macro level, it quantifies phenotypic contributions via attention-based weighting; at the micro level, Integrated Gradients reveal the genetic basis of predictions. Notably, the model successfully identified the CCT (CONSTANS, CO-like, and TOC) motif regulating photoperiodism and the SQUAMOSA (SQUAMOSA promoter binding protein) promoter for inflorescence development, confirming its ability to capture functional biological mechanisms. These results highlight the high performance and interpretability of Mul-PheG2P, showcasing its value for low-cost, large-scale screening to advance precision breeding.

Phenotype

The cold case of state transition 7 (stt7) mutants of Chlamydomonas reinhardtii, solved by whole-genome sequencing.

The process of State Transitions (ST) corresponds to an STT7 kinase-driven redistribution of the transmembrane LHCII antenna proteins between Photosystem II (PSII) and Photosystem I (PSI), which results from changes in their phosphorylation state. For the past two decades, two LHCII-kinase mutants, stt7-1 and stt7-9, have been instrumental in the study of STs in Chlamydomonas reinhardtii, the former being a null mutant for the kinase but quasi-sterile in crosses, while the latter, although fertile, has a leaky phenotype. Using long-read sequencing, this study further characterized the genetic lesions of the stt7 mutant strains through whole-genome reconstruction and de novo chromosome assembly. In addition, two new stt7 null mutants were generated, one derived by crosses from the original stt7-1 and one obtained by Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR)-associated protein 9 (Cas9) technology. This work provides a comprehensive genomic characterization of the original stt7-1 null mutant, revealing extensive chromosomal rearrangements and high levels of aneuploidy, associated with increased cell size and meiotic dysfunction. Reassessment of their physiology and genetic backgrounds highlights the need for caution in interpreting genetic information. We thus produced more reliable null mutants for the LHCII-kinase, amenable to genetic crosses for the study of STs in a variety of genetic backgrounds.

Chlamydomonas reinhardtii

Distinct functions of mammalian RAD51 paralogs in genome maintenance.

RAD51 paralogs (RAD51B, RAD51C, RAD51D, XRCC2, and XRCC3) are evolutionarily conserved essential proteins for cell survival and genome maintenance. RAD51 paralogs were originally identified to play a role in homologous recombination-mediated repair of DNA double-strand breaks (DSBs). However, investigations over the last decade have uncovered new roles of RAD51 paralogs beyond DSB repair in replication stress responses, including replication fork progression, fork stability, and its restart. Recent structural studies have not only uncovered the molecular architecture of previously known RAD51 paralog complexes but also identified novel paralog complex assemblies, providing mechanistic insights into their various genome-maintenance functions. Additionally, a role for RAD51 paralogs in resolving R-loops has been identified, and studies with cancer-associated variants suggest that RAD51 paralogs are potential determinants of cancer susceptibility and therapeutic responses. In the present review, we highlight the recently deciphered structures and novel functions of RAD51 paralog complexes and discuss the clinical and therapeutic implications.

Rad51 Recombinase

Integrated transcriptomic and metabolomic analysis of fluoride tolerance-related pathways and differentially expressed genes in silkworm strain XSKD.

XueSong KD (XSKD) silkworm strain exhibits prominent fluoride tolerance, yet the underlying molecular mechanisms of fluoride tolerance remains unclear. In the present study, fourth-instar pre-molting XSKD silkworms were used as experimental materials for integrated transcriptomic and untargeted metabolomic analyses. In total, 572 differentially expressed genes and 90 differential metabolites were screened. GO enrichment and KEGG enrichment based on the hypergeometric distribution model revealed that 13-Hydroxy-9Z,11E-octadecadienoic acid (13-(S)-HODE) acts as the core differential metabolite, which is significantly enriched in the linoleic acid metabolism pathway. Within this pathway, LOC101737302 and CYP338A1 display opposite expression trends and show correlations with pathway metabolites. Based on multi-omics data, this study preliminarily characterizes the lipid metabolic response under fluoride stress, providing omics dataset support for further in-depth exploration of the molecular mechanism of fluoride tolerance in silkworms.

Animals

Gloeotrichia echinulata genomes from the United States are nontoxigenic and likely geosmin producers.

Six Gloeotrichia echinulata genomes derived from planktonic harmful algal blooms (HABs) with similar colonial morphology have been sequenced from lakes in the west and northeast regions of USA, four of them to completion. The c. 7 Mbp genomes exhibit a high level of conservation, with 98-99% pairwise genome-wide average nucleotide identity and high levels of synteny, representing a single species cluster. We observed strong conservation of gene clusters responsible for the synthesis of the secondary metabolites and bioactive peptides that are characteristic of HAB-forming cyanobacteria. All six G. echinulata genomes lack genes for the synthesis of classic cyanotoxins, including microcystin, but possess genes responsible for the synthesis of the taste and odor compound geosmin. Interestingly, the geoA geosmin synthase gene in three genomes is homologous to other cyanobacterial geoA genes, while the other three geoA genes are related to actinomyces geoA. Phylogenomic analysis places the G. echinulata genomes within a clade of benthic Nostocales, reflecting an ecological niche featuring extensive growth on the sediment surface before colonies disperse into the epilimnion for planktonic growth. We identify genes conserved in all six genomes that could represent physiological adaptations supporting active growth on sediments and pelagic recruitment independent of wind-driven mixing: phycoerythrin light harvesting complexes for optimal photosynthesis at depth; gliding motility to access patchy nutrient distributions; and gas vesicles with relatively small GvpC proteins that predict resistance to higher hydrostatic pressure. The strong genomic similarity across geographically distant populations suggests that G. echinulata in the United States is a tightly related non-toxigenic species group with predictable properties relevant to public health and drinking water management.

Cyanobacteria

Efficient homologous replacement and deletion of large genomic fragments through template-jumping prime editing in rice.

Homologous replacement of genomic sequences with large DNA fragments (> 100 bp) holds great potential for crop breeding, yet an efficient method to achieve such edits is lacking in plants. Here, in rice, we developed template-jumping prime editing (TJ-PE), a recently reported PE strategy for large targeted insertion, as an efficient tool for homologous replacement with DNA fragments ranging from dozens to hundreds of base pairs, and using TJ-PE, we replaced genomic fragments of up to 340 bp with homologous fragments of the same length. In addition, our TJ-PE tool also enabled precise deletion of 944- to 2024-bp fragments in rice, with efficiencies of up to 34.6% for c. 2000-bp precise deletions. Collectively, this study expands the editing scope of PE in rice and establishes TJ-PE as a generalist tool for precise deletion and replacement of large DNA fragments.

Oryza