Search PubMedSearch

SEARCH · Search PubMed

Results for “genome skims”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

379 records · Page 2Linked to original sources

The cold case of state transition 7 (stt7) mutants of Chlamydomonas reinhardtii, solved by whole-genome sequencing.

The process of State Transitions (ST) corresponds to an STT7 kinase-driven redistribution of the transmembrane LHCII antenna proteins between Photosystem II (PSII) and Photosystem I (PSI), which results from changes in their phosphorylation state. For the past two decades, two LHCII-kinase mutants, stt7-1 and stt7-9, have been instrumental in the study of STs in Chlamydomonas reinhardtii, the former being a null mutant for the kinase but quasi-sterile in crosses, while the latter, although fertile, has a leaky phenotype. Using long-read sequencing, this study further characterized the genetic lesions of the stt7 mutant strains through whole-genome reconstruction and de novo chromosome assembly. In addition, two new stt7 null mutants were generated, one derived by crosses from the original stt7-1 and one obtained by Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR)-associated protein 9 (Cas9) technology. This work provides a comprehensive genomic characterization of the original stt7-1 null mutant, revealing extensive chromosomal rearrangements and high levels of aneuploidy, associated with increased cell size and meiotic dysfunction. Reassessment of their physiology and genetic backgrounds highlights the need for caution in interpreting genetic information. We thus produced more reliable null mutants for the LHCII-kinase, amenable to genetic crosses for the study of STs in a variety of genetic backgrounds.

Chlamydomonas reinhardtii

Comparative genomics and full-length transcriptome profiling of wing morphs in Tetrix grossus (Orthoptera: Tetrigidae).

Wing polymorphism represents a paradigmatic dispersal-reproduction trade-off, yet its molecular basis remains uncharacterised in the phylogenetically distant pygmy grasshoppers (Tetrigidae). Here we integrate comparative genomics across ten orthopteran species with full-length transcriptomics of long-winged (FL) and short-winged (FS) Tetrix grossus. OrthoFinder recovered 118 orthogroups specific to T. grossus. Against a backdrop of pronounced gene-family contraction (36 expansions versus 222 contractions; net -186, mirrored at the ancestral Tetrix node, +37/-140), we identified an ancestral, Tetrix-specific expansion of hormone-regulation (12 genes; fold enrichment 7.93) and lipid/carbohydrate-metabolic families organised into syntenic clusters, alongside 513 positively selected genes enriched for integrin-mediated cell adhesion (6 genes), a process relevant to epithelial and appendage morphogenesis. Full-length transcriptomics of one long-winged (FL) and one short-winged (FS) adult female detected 7530 (FL) and 7515 (FS) expressed genes, with 794 FL- and 776 FS-restricted transcriptome-derived SNP-associated genes. The FL morph was enriched for an EGFR/Ras-Rho developmental-patterning axis and neuromuscular flight genes, whereas the FS morph was enriched for insulin/peptide-hormone response and growth-regulatory loci. Overall, we present genomic resources and testable hypotheses concerning the evolution and regulation of wing morphs in Tetrigidae rather than a validated genetic architecture of wing-morph determination.

Animals

Genome-wide characterization of heat shock protein genes reveals thermal stress-responsive candidates in Litopenaeus vannamei.

Heat shock proteins (HSPs) are conserved molecular chaperones involved in protein folding, refolding, aggregation prevention, and degradation of damaged proteins. However, the genomic organization and thermal responsiveness of HSP genes in the Pacific white shrimp (Litopenaeus vannamei) remain incompletely understood. Here, we performed a genome-wide analysis of the HSP gene family and examined its phylogenetic relationships, structural features, duplication patterns, sequence variation, interaction networks, and transcriptional responses to acute heat stress. A total of 34 HSP genes were identified and classified into the HSP90, HSP70, HSP40/DNAJ, HSP60, and small HSP families. Phylogenetic, motif, gene structure, synteny, and subcellular localization analyses revealed evolutionary conservation and structural diversification among family members. Three duplicated gene pairs were identified, comprising two segmental duplications and one tandem duplication. All pairs exhibited Ka/Ks ratios below 1, consistent with purifying selection of varying strength. Sequence analysis identified 295 nonsynonymous single-nucleotide polymorphisms, of which 12 were consistently predicted to be deleterious by multiple algorithms. Protein-protein interaction analysis indicated enrichment of protein-folding and cellular stress-response functions. RT-qPCR analysis showed significant induction of HSPA4, HSP90AA1, TRAP1, BiP, and DNAJA1 after 6, 12, and 24 h of exposure to 34 °C, whereas DNAJC3 was significantly induced only at 12 h. All six genes reached their highest transcript abundance at 12 h. These findings may provide a genomic framework for HSP genes in L. vannamei and identify candidate genes and variants associated with thermal stress responses.

Animals

Gloeotrichia echinulata genomes from the United States are nontoxigenic and likely geosmin producers.

Six Gloeotrichia echinulata genomes derived from planktonic harmful algal blooms (HABs) with similar colonial morphology have been sequenced from lakes in the west and northeast regions of USA, four of them to completion. The c. 7 Mbp genomes exhibit a high level of conservation, with 98-99% pairwise genome-wide average nucleotide identity and high levels of synteny, representing a single species cluster. We observed strong conservation of gene clusters responsible for the synthesis of the secondary metabolites and bioactive peptides that are characteristic of HAB-forming cyanobacteria. All six G. echinulata genomes lack genes for the synthesis of classic cyanotoxins, including microcystin, but possess genes responsible for the synthesis of the taste and odor compound geosmin. Interestingly, the geoA geosmin synthase gene in three genomes is homologous to other cyanobacterial geoA genes, while the other three geoA genes are related to actinomyces geoA. Phylogenomic analysis places the G. echinulata genomes within a clade of benthic Nostocales, reflecting an ecological niche featuring extensive growth on the sediment surface before colonies disperse into the epilimnion for planktonic growth. We identify genes conserved in all six genomes that could represent physiological adaptations supporting active growth on sediments and pelagic recruitment independent of wind-driven mixing: phycoerythrin light harvesting complexes for optimal photosynthesis at depth; gliding motility to access patchy nutrient distributions; and gas vesicles with relatively small GvpC proteins that predict resistance to higher hydrostatic pressure. The strong genomic similarity across geographically distant populations suggests that G. echinulata in the United States is a tightly related non-toxigenic species group with predictable properties relevant to public health and drinking water management.

Cyanobacteria

Genomic determinants underlying biogenic amine detoxification phenotypes in food-associated lactic acid bacteria: Mechanism, evolutionary origin, and relevance to fermented food safety.

Biogenic amines (BAs) are toxic metabolites that accumulate in fermented foods and pose significant food safety concerns. Although several lactic acid bacteria (LAB) have previously been reported to exhibit strain-specific BA-degrading phenotypes, the genetic determinants underlying these activities have remained largely uncharacterized. Here, we analyzed 8251 LAB genomes to validate BA-degrading phenotypes. We predicted five BA-associated genes, including two direct biogenic amine-degrading genes (BADGs), mco and patA, and three polyamine-modifying genes (PMGs), speG, paiA, and bltD. Among BADGs, mco was broadly distributed across LAB and strongly enriched across food-associated niches. patA, organized within a conserved potD-glnB-potABC-patA cassette, is a putative, functionally distinct BADG in LAB, revealing a nitrogen-responsive polyamine uptake-catabolism module. Phylogenomics, phylogenetic reconciliation, and synteny analysis established that all five genes entered the LAB through episodic horizontal gene transfer followed by lineage-specific fixation. GC compositional bias and mobile genetic element association further corroborated the horizontal origin of the two BADGs. Structural analysis confirmed the conservation of catalytic core residues of BADGs across LAB, indicating strong purifying selection. Phenotype-to-genotype correlation with experimentally reported LAB suggested mco as a reliable genomic predictor of degrading phenotype. Integration of degradation and biosynthetic profiles predicted multiple LAB species capable of both synthesizing and degrading BA, along with 1823 genomes with degradation potential but lacking detectable BA biosynthesis genes. This study provides the first large-scale genome framework linking BA-degrading phenotypes with their genetic determinants in LAB and offers a rational basis for selecting BA-detoxifying strains for fermented food applications.

Biogenic Amines

Efficient homologous replacement and deletion of large genomic fragments through template-jumping prime editing in rice.

Homologous replacement of genomic sequences with large DNA fragments (> 100 bp) holds great potential for crop breeding, yet an efficient method to achieve such edits is lacking in plants. Here, in rice, we developed template-jumping prime editing (TJ-PE), a recently reported PE strategy for large targeted insertion, as an efficient tool for homologous replacement with DNA fragments ranging from dozens to hundreds of base pairs, and using TJ-PE, we replaced genomic fragments of up to 340 bp with homologous fragments of the same length. In addition, our TJ-PE tool also enabled precise deletion of 944- to 2024-bp fragments in rice, with efficiencies of up to 34.6% for c. 2000-bp precise deletions. Collectively, this study expands the editing scope of PE in rice and establishes TJ-PE as a generalist tool for precise deletion and replacement of large DNA fragments.

Oryza

Ramu stunt virus genome reveals previously unreported segments and nucleocapsid domain duplication in Mechlorovirus.

Ramu stunt virus (RmSV), a member of the genus Mechlorovirus within the family Phenuiviridae, was previously described as a six-segmented RNA virus infecting sugarcane. In this study, we re-examined type material and additional isolates using high-throughput sequencing and RT-PCR validation, revealing that RmSV possesses a nine-segmented genome, making it the largest reported in the Phenuiviridae. This expanded architecture includes duplicated RNA segments (RNA 2a and RNA 2b) encoding nucleocapsid-like proteins and two novel segments (RNA 7 and RNA 8). Comparative analysis showed that RNA 2a and 2b share about 84% amino acid identity, while RNA 5 encodes a third nucleocapsid homolog, indicating unprecedented domain redundancy. Structural modeling confirmed that all three nucleocapsid proteins maintain a conserved fold despite low sequence identity, with electrostatic mapping suggesting differential RNA-binding potential. Additionally, RNA 6 encodes a hypothetical protein structurally similar to the rice stripe virus disease-specific S-protein, implicating a role in symptom development. Transcript abundance analysis revealed RNA 6 as the most highly expressed segment across isolates. These findings revise the genomic composition of RmSV, highlight mechanisms of genome plasticity and adaptive evolution in plant-infecting bunyaviruses, and underscore practical implications for diagnostic assay design, resistance breeding, and biosecurity surveillance.

Genome, Viral

Genomic and Phenotypic Characterization of Two Novel Enterobacter Phages With EDTA-Enhanced Antibiofilm Activity.

Multidrug-resistant members of the Enterobacter cloacae complex (ECC) are increasingly linked to difficult-to-treat infections and biofilm-mediated antimicrobial tolerance. Here, two lytic phages, vB_EhoIP_HHH and vB_EluM_RZH, displaying podovirus-like and myovirus-like morphology, respectively, were isolated from the River Chelt. HHH has a 39,582 bp genome (51.2% GC, 63 ORFs), while RZH has a 174,197 bp genome (39.4% GC, 314 ORFs), with neither genome carrying antimicrobial resistance, virulence or lysogeny-associated genes. VIRIDIC and VICTOR analyses placed HHH within Kayfunavirus and RZH within Karamvirus, supporting their classification as distinct species. Both phages demonstrated rapid adsorption, short latent periods and stability across physiological pH and temperature ranges. A phage cocktail targeting MDR ECC strain was evaluated with EDTA against established biofilms. Crystal violet assays showed the greatest biomass reduction at MOI 10 with 0.5-0.75 mM EDTA. Bliss independence analysis revealed localized synergy within this window but significant overall antagonism at higher EDTA concentrations. CFU enumeration confirmed greater activity against 24 h than 48 h biofilms. The optimized combination also reduced recoverable bacteria in a fibroblast infection model while maintaining low LDH release. These findings identify two novel lytic Enterobacter phages and support a narrow EDTA concentration window for enhanced phage-mediated antibiofilm activity.

Biofilms

Cost-Effectiveness and the Economics of Genomic Testing and Molecularly Matched Therapies.

Cost-effectiveness analysis of precision oncology can help guide value-driven care. Next-generation sequencing is increasingly cost-efficient over single gene testing because diagnostic algorithms require multiple individual gene tests to determine biomarker status. Matched targeted therapy is often not cost-effective due to the high cost associated with drug treatment. However, genomic profiling can promote cost-effective care by identifying patients who are unlikely to benefit from therapy. Additional applications of genomic profiling such as universal testing for hereditary cancer syndromes and germline testing in patients with cancer may represent cost-effective approaches compared with traditional history-based diagnostic methods.

Humans

Whole genome sequencing of unusual Hepatitis C virus subtypes and drug resistance analysis during direct-acting antiviral therapy in India.

INTRODUCTION AND OBJECTIVES: Pangenotypic direct-acting antivirals (DAA) are effective against highly prevalent Hepatitis C virus (HCV) subtypes, but have been clinically validated almost exclusively in high-income countries. Unusual HCV subtypes may carry natural polymorphisms, potentially impacting DAA susceptibility. We conducted full-genome characterization and resistance analysis of unusual HCV subtypes in patients receiving DAA treatment. PATIENTS AND METHODS: In this prospective hospital-based study, eligible patients were screened for anti-HCV antibodies and active infection was confirmed by diagnostic 5'NCR-based HCV RNA detection. Genotyping was performed by core region sequencing, and viral load quantified by real-time PCR. For whole genome sequencing, multiplex primers were designed using alignments of global reference sequences. Sequencing was carried out using the Oxford Nanopore Technology platform. Phylogenetic analysis used multiple sequence alignment and the HCV-GLUE resource for resistance-associated substitution (RAS) analysis. RESULTS: Predominant genotype was genotype 3 in 64.3% (n = 45); genotype 6 in 21.4% (n = 15); and genotype 1 in 14.2% (n = 10). Unusual HCV subtype 6xa was detected in two patients and showed no NS5A resistance mutations. One genotype 3b patient relapsed at 24 weeks post-DAA treatment completion and carried NS5A resistance-associated substitutions 30 K and 31 M both at baseline and at relapse, conferring high-level resistance to NS5A inhibitors. CONCLUSION: This is the first report from India of whole genome sequencing of HCV subtype 6xa. The identification of NS5A resistance mutations in the 3b relapse case underscores challenges for global HCV elimination strategies.

Humans

Revealing the Shared Genetic Architecture of Metabolic Dysfunction-Associated Steatotic Liver Disease-Related Traits Through Genomic Structural Equation Modeling.

Although individual traits related to metabolic dysfunction-associated steatotic liver disease (MASLD) have been investigated through large-scale genome-wide association studies (GWASs), the shared genetic susceptibility across these traits remains unclear. We therefore conducted a multivariate GWAS of key MASLD-related traits to elucidate their common genetic architecture. We applied genomic structural equation modeling to model a latent genetic factor (MASLD-F) underlying genetically correlated MASLD-related traits, leveraging their GWAS-derived genetic correlations. We then performed functional annotations, including fine-mapping, transcriptome-wide association study, and cell- and tissue-type-specific enrichment analyses, and conducted Mendelian randomization analyses to identify modifiable risk factors. Our multivariate MASLD-F GWAS identified 50 independent variants across 48 genomic loci. Transcriptomic imputation identified several MASLD-F-associated genes, including ARNTL, NPC1, BTBD10, VDAC2, TSKU, SFMBT1, and ABHD17C. We observed significant enrichment of MASLD-F-related genetic signals predominantly in brain tissues, pancreatic islets, and the adrenal gland. Additionally, six modifiable risk factors and four modifiable protective factors for MASLD-F were identified. These findings reveal a complex shared genetic architecture underlying MASLD components, thereby expanding our understanding of disease pathogenesis and providing novel insights for precision medicine and public health interventions.

Humans

Exploring the mechanism of aroma production in fermented cherry juice by L. brevis LD1.0600 using flavomics and whole genome analysis.

This study focused on L.brevis LD1.0600 with excellent fermentation traits: it analyzed genome-wide key regulatory genes for micro-metabolites, combined with fermented cherry juice flavor metabolomics data, and used machine learning to explore correlations between gene regulation, metabolite production, and flavor formation. The SVM model screened and verified fermented cherry juice VOCs; through OAV and flavor wheel analysis, LD1.0600 emerged as the top-performing strain, with a sweet, fruity dominant aroma. Key aroma-active components (OAV > 100) included 2-methoxy-4-vinylphenol, benzaldehyde, 2-methyl-butanoic acid and hexanoic acid, and 2-methoxy-4-vinylphenol and hexanoic acid elevated by LD1.0600-regulated genes (Chrom1-001884, Chrom1-000925, fabF and Chrom1-000199). At the same time, through research, a "strain screening-SVM screening of DVCs-OAV screening of key aroma components-whole genome sequencing of flavor regulatory genes" system was established. This system can not only be applied to the screen fermentation strains, but also can be extended to the application of other fermentation products.

Fermentation

Introgression shapes the genomic conflict landscape of Malus, providing evidence for a reticulate backbone in a woody crop lineage.

Phylogenomic discordance is widespread across plants, but its evolutionary significance is often obscured when conflict is treated primarily as analytical noise rather than as evidence of underlying processes. In woody lineages in particular, incomplete lineage sorting, introgression, and genome duplication can interact over long timescales to produce complex genomic histories that are not adequately summarized by a strictly bifurcating tree. Here, we use Malus as a model woody genus to investigate how these processes structure conflict across a genus-scale, accession-based phylogenomic framework. Using broad taxon sampling, hundreds of nuclear loci, plastid genomes, and genome-wide SNP summaries, we reconstruct a robust nuclear backbone for sampled Malus lineages and evaluate where discordance is concentrated and which processes best explain it. Nuclear analyses resolve eight major clades, whereas conflict is non-random and localized to recurrent hotspots rather than evenly distributed across the tree. Cytonuclear discordance is similarly concentrated, especially around Clade H, represented by sampled accessions of M. tschonoskii, where localized plastid-nuclear disagreement is consistent with candidate plastid capture or organellar introgression. Multiple complementary analyses further indicate that the strongest conflict is not explained by ILS alone, but instead reflects lineage-structured introgression, while polyploid complexes represent additional localized sources of evolutionary complexity. Together, these results provide evidence for a reticulate genomic backbone in Malus and show how integrating nuclear, plastid, and genome-wide conflict analyses can help distinguish background discordance from process-specific signals in woody plant radiations. Several lineage-level reticulation hypotheses identified here should now be tested with broader population-level sampling and curated reference accessions.

Malus

Shared genetic architecture and cellular convergence between female reproductive disorders and pulmonary function: a genome-wide cross-trait analysis.

Female reproductive disorders (FRDs), including polycystic ovary syndrome, endometriosis, uterine leiomyomata, and infertility, have been epidemiologically associated with impaired pulmonary function. However, it remains unclear whether this cross-organ link reflects shared genetic etiology and, if so, which cellular mechanisms mediate it. We performed a systematic genome-wide cross-trait analysis of three FRDs and lung function traits (FEV₁, FVC, FEV₁/FVC) using GWAS summary statistics from individuals of European ancestry, integrating genetic correlation, bidirectional causal inference, pleiotropy mapping, and single-cell enrichment analyses. We identified significant negative genetic correlations between FRDs and lung volume traits, most prominently for FVC (rg range: - 0.077 to - 0.178). Bidirectional causal analyses indicated that FRDs have a detrimental effect on lung volume, with higher FRD genetic liability associated with reduced lung volume. Cross-trait meta-analysis identified 17 pleiotropic variants across 11 loci, with the 19q13.2 (LTBP4) and 12q13.13 (HOXC6/HOXC9) loci showing strong evidence of shared causal variants. Critically, single-cell analyses revealed that shared genetic risk converged on mesenchymal lineages across organs, specifically alveolar adventitial fibroblasts in the lung and stromal/smooth muscle cells in the endometrium. Transcriptome-wide analyses further nominated the estrogen-responsive gene RERG as a convergent gene linking these conditions with lung function. Our study revealed a shared genetic architecture between female reproductive disorders and lung function traits, providing a basis for further mechanistic investigations and potential clinical evaluation. Furthermore, our findings suggest that shared fibroproliferative and hormone-responsive pathways may offer insights into the biological mechanisms underlying these conditions.

Female

Genome-wide identification, structural characterization, and evolutionary analysis of growth-related gene families in African catfish (Clarias gariepinus).

The somatotropic axis encompassing growth hormone (GH), insulin-like growth factor (IGF), myostatin (MSTN), and prolactin (PRL) signalling cascades is the master regulator of somatic growth, metabolism, and development in vertebrates. African catfish (Clarias gariepinus), a commercially pivotal aquaculture species, now possesses a chromosome-level reference genome (CGAR_prim_01v2); however, a systematic, genome-wide characterization spanning all five interconnected growth-related gene families has not previously been undertaken in this species. Here, we identified and characterized 15 growth-related genes spanning gh1, ghra, ghrb, Igf1, Igf2a, Igf2b, igf1ra, Igf1rb, Igf2r, Mstna, Mstnb, prl, prlra, prlrb, and smtlb distributed across 13 chromosomes. Complete one-to-one orthology with zebrafish confirmed strong dosage-balance conservation across >120 million years of teleost divergence. Physicochemical analysis resolved a clear biochemical dichotomy between compact, basic secreted ligands (19.88-45.81 kDa; pI up to 10.02) and large, acidic, heavily glycosylated membrane receptors (56.82-270.80 kDa; pI 4.85-5.97). Phylogenetic analysis confirmed 3R whole-genome duplication origins for all paralog pairs, while synteny analysis revealed a disruption of the ancestral gh1-prl chromosomal block in C. gariepinus, a finding that warrants further comparative and functional investigation. This genomic atlas provides the sequence and structural information including exon-intron boundaries, domain architecture, and chromosomal coordinates needed as a prerequisite for future marker-assisted selection and CRISPR-based myostatin-editing efforts in African catfish aquaculture, though translation into applied breeding outcomes will require subsequent functional and expression studies.

Animals