Search PubMedSearch

SEARCH · Search PubMed

Results for “Genome-guided transcriptome”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

740 records · Page 15Linked to original sources

Distinct functions of mammalian RAD51 paralogs in genome maintenance.

RAD51 paralogs (RAD51B, RAD51C, RAD51D, XRCC2, and XRCC3) are evolutionarily conserved essential proteins for cell survival and genome maintenance. RAD51 paralogs were originally identified to play a role in homologous recombination-mediated repair of DNA double-strand breaks (DSBs). However, investigations over the last decade have uncovered new roles of RAD51 paralogs beyond DSB repair in replication stress responses, including replication fork progression, fork stability, and its restart. Recent structural studies have not only uncovered the molecular architecture of previously known RAD51 paralog complexes but also identified novel paralog complex assemblies, providing mechanistic insights into their various genome-maintenance functions. Additionally, a role for RAD51 paralogs in resolving R-loops has been identified, and studies with cancer-associated variants suggest that RAD51 paralogs are potential determinants of cancer susceptibility and therapeutic responses. In the present review, we highlight the recently deciphered structures and novel functions of RAD51 paralog complexes and discuss the clinical and therapeutic implications.

Rad51 Recombinase

Genome-wide characterization of heat shock protein genes reveals thermal stress-responsive candidates in Litopenaeus vannamei.

Heat shock proteins (HSPs) are conserved molecular chaperones involved in protein folding, refolding, aggregation prevention, and degradation of damaged proteins. However, the genomic organization and thermal responsiveness of HSP genes in the Pacific white shrimp (Litopenaeus vannamei) remain incompletely understood. Here, we performed a genome-wide analysis of the HSP gene family and examined its phylogenetic relationships, structural features, duplication patterns, sequence variation, interaction networks, and transcriptional responses to acute heat stress. A total of 34 HSP genes were identified and classified into the HSP90, HSP70, HSP40/DNAJ, HSP60, and small HSP families. Phylogenetic, motif, gene structure, synteny, and subcellular localization analyses revealed evolutionary conservation and structural diversification among family members. Three duplicated gene pairs were identified, comprising two segmental duplications and one tandem duplication. All pairs exhibited Ka/Ks ratios below 1, consistent with purifying selection of varying strength. Sequence analysis identified 295 nonsynonymous single-nucleotide polymorphisms, of which 12 were consistently predicted to be deleterious by multiple algorithms. Protein-protein interaction analysis indicated enrichment of protein-folding and cellular stress-response functions. RT-qPCR analysis showed significant induction of HSPA4, HSP90AA1, TRAP1, BiP, and DNAJA1 after 6, 12, and 24 h of exposure to 34 °C, whereas DNAJC3 was significantly induced only at 12 h. All six genes reached their highest transcript abundance at 12 h. These findings may provide a genomic framework for HSP genes in L. vannamei and identify candidate genes and variants associated with thermal stress responses.

Animals

Gloeotrichia echinulata genomes from the United States are nontoxigenic and likely geosmin producers.

Six Gloeotrichia echinulata genomes derived from planktonic harmful algal blooms (HABs) with similar colonial morphology have been sequenced from lakes in the west and northeast regions of USA, four of them to completion. The c. 7 Mbp genomes exhibit a high level of conservation, with 98-99% pairwise genome-wide average nucleotide identity and high levels of synteny, representing a single species cluster. We observed strong conservation of gene clusters responsible for the synthesis of the secondary metabolites and bioactive peptides that are characteristic of HAB-forming cyanobacteria. All six G. echinulata genomes lack genes for the synthesis of classic cyanotoxins, including microcystin, but possess genes responsible for the synthesis of the taste and odor compound geosmin. Interestingly, the geoA geosmin synthase gene in three genomes is homologous to other cyanobacterial geoA genes, while the other three geoA genes are related to actinomyces geoA. Phylogenomic analysis places the G. echinulata genomes within a clade of benthic Nostocales, reflecting an ecological niche featuring extensive growth on the sediment surface before colonies disperse into the epilimnion for planktonic growth. We identify genes conserved in all six genomes that could represent physiological adaptations supporting active growth on sediments and pelagic recruitment independent of wind-driven mixing: phycoerythrin light harvesting complexes for optimal photosynthesis at depth; gliding motility to access patchy nutrient distributions; and gas vesicles with relatively small GvpC proteins that predict resistance to higher hydrostatic pressure. The strong genomic similarity across geographically distant populations suggests that G. echinulata in the United States is a tightly related non-toxigenic species group with predictable properties relevant to public health and drinking water management.

Cyanobacteria

Genomic determinants underlying biogenic amine detoxification phenotypes in food-associated lactic acid bacteria: Mechanism, evolutionary origin, and relevance to fermented food safety.

Biogenic amines (BAs) are toxic metabolites that accumulate in fermented foods and pose significant food safety concerns. Although several lactic acid bacteria (LAB) have previously been reported to exhibit strain-specific BA-degrading phenotypes, the genetic determinants underlying these activities have remained largely uncharacterized. Here, we analyzed 8251 LAB genomes to validate BA-degrading phenotypes. We predicted five BA-associated genes, including two direct biogenic amine-degrading genes (BADGs), mco and patA, and three polyamine-modifying genes (PMGs), speG, paiA, and bltD. Among BADGs, mco was broadly distributed across LAB and strongly enriched across food-associated niches. patA, organized within a conserved potD-glnB-potABC-patA cassette, is a putative, functionally distinct BADG in LAB, revealing a nitrogen-responsive polyamine uptake-catabolism module. Phylogenomics, phylogenetic reconciliation, and synteny analysis established that all five genes entered the LAB through episodic horizontal gene transfer followed by lineage-specific fixation. GC compositional bias and mobile genetic element association further corroborated the horizontal origin of the two BADGs. Structural analysis confirmed the conservation of catalytic core residues of BADGs across LAB, indicating strong purifying selection. Phenotype-to-genotype correlation with experimentally reported LAB suggested mco as a reliable genomic predictor of degrading phenotype. Integration of degradation and biosynthetic profiles predicted multiple LAB species capable of both synthesizing and degrading BA, along with 1823 genomes with degradation potential but lacking detectable BA biosynthesis genes. This study provides the first large-scale genome framework linking BA-degrading phenotypes with their genetic determinants in LAB and offers a rational basis for selecting BA-detoxifying strains for fermented food applications.

Biogenic Amines

Efficient homologous replacement and deletion of large genomic fragments through template-jumping prime editing in rice.

Homologous replacement of genomic sequences with large DNA fragments (> 100 bp) holds great potential for crop breeding, yet an efficient method to achieve such edits is lacking in plants. Here, in rice, we developed template-jumping prime editing (TJ-PE), a recently reported PE strategy for large targeted insertion, as an efficient tool for homologous replacement with DNA fragments ranging from dozens to hundreds of base pairs, and using TJ-PE, we replaced genomic fragments of up to 340 bp with homologous fragments of the same length. In addition, our TJ-PE tool also enabled precise deletion of 944- to 2024-bp fragments in rice, with efficiencies of up to 34.6% for c. 2000-bp precise deletions. Collectively, this study expands the editing scope of PE in rice and establishes TJ-PE as a generalist tool for precise deletion and replacement of large DNA fragments.

Oryza

Reference genome of the Californian trapdoor spider Aptostichus stephencolberti Bond 2008 (Araneae: Mygalomorphae: Euctenizidae).

We present a reference genome assembly for the trapdoor spider Aptostichus stephencolberti. This species, described in 2008, is endemic to the highly fragmented coastal dune habitats of Northern California from Monterey to the San Francisco Bay Area. Trapdoor spiders are ideal taxa for landscape scale genomic studies owing to their extreme site fidelity and limited dispersal capabilities; these same characteristics make them prone to extinction. Genomic studies of species like A. stephencolberti can reveal novel areas of endemism and high conservation value that may not be evident in species with wider ranges and greater dispersal capabilities. As part of the California Conservation Genomics Project, we constructed the A. stephencolberti reference genome from high quality long-read sequences, scaffolded with proximity ligation Omni-C data. The primary assembly comprises 551 scaffolds spanning 3.63 Gbp, a scaffold N50 of 62.2 Mbp and BUSCO completeness of 95.6%. We estimate 52 chromosomes yet find no (TTAGG)n telomer repeats. Expanding the telomeric repeat search finds an ancestral loss of the repeat from all spiders. Automated annotation using the NCBI refseq pipeline and RNAseq data from whole adults finds 14,067 genes with a BUSCO annotation completeness of 95.56%. Repeat annotation identified 77% of the genome to be interspersed repeats. This resource, the first for family Euctenizidae will facilitate future study and resulting conservation actions of A. stephencolberti and other Aptostichus sp. populations associated with the rapidly changing California coastal dune ecosystem.

Aptostichus stephencolberti

Leveraging traveller genomics for LMIC diarrhoeal disease management.

Diarrhoeal pathogens impose a substantial global health burden, disproportionately affecting low- and middle-income countries (LMICs). However, in these settings, health-seeking behaviours, suboptimal microbiological capacity, and challenges in establishing genomics capacity constrain effective surveillance, including surveillance of antimicrobial resistance (AMR). In contrast, high-income countries routinely generate and share large volumes of diarrhoeal pathogen genomes through established systems, with a significant proportion originating from travellers returning from LMICs. These data reveal strong geographical structuring of lineages and clinically relevant AMR patterns, demonstrating untapped potential to support improvements in geographically granulated surveillance to support antimicrobial treatment recommendations. In this opinion article, we outline the potential to integrate traveller-derived microbial genomic data into LMIC public health decision-making and highlight the scientific, ethical, practical, and governance considerations for implementation.

antimicrobial resistance

Cytonuclear conflict and reticulate evolution in the Morelloid clade (Solanum, Solanaceae): Insights from genome skimming and network Phylogenomics.

The Morelloid clade (black nightshades) is one of the most strongly supported clades within the megadiverse Solanum genus. It comprises 76 globally distributed, non-spiny herbaceous and suffrutescent species. While often erroneously considered poisonous weeds, several species are economically important as orphan crops. The clade is closely related to tomato and potato but, due to a lack of focused breeding efforts, remains a putative reservoir of genetic diversity for crop improvement. Despite this potential, we lack fundamental knowledge on the evolution of the Morelloid clade. The group includes polyploid species with unknown parental origins-likely reflecting reticulate processes such as hybridization, introgression, and associated backcrossing events. Prior analyses have been unable to disentangle these processes, leaving the mechanisms underlying reticulate evolution in the Morelloid clade poorly understood. Here, we use genome skimming to produce a well-supported maximum likelihood plastid phylogeny from complete circularized plastomes and a coalescent-based species tree from combined Angiosperms353 and conserved ortholog set nuclear markers. Our dataset, composed of previously published data and deep genome skimming from herbarium samples, spans 26 Morelloid species. To investigate phylogenetic discordance, we used a nuclear phylogenetic network, multispecies coalescent simulations, a fused rooted nuclear chloroplast tree, and quantification of nuclear gene tree concordance. We show that incongruence between nuclear and plastid trees is pervasive and cannot be explained by incomplete lineage sorting alone. Instead, our results demonstrate that events consistent with repeated chloroplast capture have shaped the reticulate evolutionary history of the clade, especially among African polyploid and Pan-American diploid lineages.

Phylogeny

Ramu stunt virus genome reveals previously unreported segments and nucleocapsid domain duplication in Mechlorovirus.

Ramu stunt virus (RmSV), a member of the genus Mechlorovirus within the family Phenuiviridae, was previously described as a six-segmented RNA virus infecting sugarcane. In this study, we re-examined type material and additional isolates using high-throughput sequencing and RT-PCR validation, revealing that RmSV possesses a nine-segmented genome, making it the largest reported in the Phenuiviridae. This expanded architecture includes duplicated RNA segments (RNA 2a and RNA 2b) encoding nucleocapsid-like proteins and two novel segments (RNA 7 and RNA 8). Comparative analysis showed that RNA 2a and 2b share about 84% amino acid identity, while RNA 5 encodes a third nucleocapsid homolog, indicating unprecedented domain redundancy. Structural modeling confirmed that all three nucleocapsid proteins maintain a conserved fold despite low sequence identity, with electrostatic mapping suggesting differential RNA-binding potential. Additionally, RNA 6 encodes a hypothetical protein structurally similar to the rice stripe virus disease-specific S-protein, implicating a role in symptom development. Transcript abundance analysis revealed RNA 6 as the most highly expressed segment across isolates. These findings revise the genomic composition of RmSV, highlight mechanisms of genome plasticity and adaptive evolution in plant-infecting bunyaviruses, and underscore practical implications for diagnostic assay design, resistance breeding, and biosecurity surveillance.

Genome, Viral

Genomic and Phenotypic Characterization of Two Novel Enterobacter Phages With EDTA-Enhanced Antibiofilm Activity.

Multidrug-resistant members of the Enterobacter cloacae complex (ECC) are increasingly linked to difficult-to-treat infections and biofilm-mediated antimicrobial tolerance. Here, two lytic phages, vB_EhoIP_HHH and vB_EluM_RZH, displaying podovirus-like and myovirus-like morphology, respectively, were isolated from the River Chelt. HHH has a 39,582 bp genome (51.2% GC, 63 ORFs), while RZH has a 174,197 bp genome (39.4% GC, 314 ORFs), with neither genome carrying antimicrobial resistance, virulence or lysogeny-associated genes. VIRIDIC and VICTOR analyses placed HHH within Kayfunavirus and RZH within Karamvirus, supporting their classification as distinct species. Both phages demonstrated rapid adsorption, short latent periods and stability across physiological pH and temperature ranges. A phage cocktail targeting MDR ECC strain was evaluated with EDTA against established biofilms. Crystal violet assays showed the greatest biomass reduction at MOI 10 with 0.5-0.75 mM EDTA. Bliss independence analysis revealed localized synergy within this window but significant overall antagonism at higher EDTA concentrations. CFU enumeration confirmed greater activity against 24 h than 48 h biofilms. The optimized combination also reduced recoverable bacteria in a fibroblast infection model while maintaining low LDH release. These findings identify two novel lytic Enterobacter phages and support a narrow EDTA concentration window for enhanced phage-mediated antibiofilm activity.

Biofilms

Whole genome sequencing of unusual Hepatitis C virus subtypes and drug resistance analysis during direct-acting antiviral therapy in India.

INTRODUCTION AND OBJECTIVES: Pangenotypic direct-acting antivirals (DAA) are effective against highly prevalent Hepatitis C virus (HCV) subtypes, but have been clinically validated almost exclusively in high-income countries. Unusual HCV subtypes may carry natural polymorphisms, potentially impacting DAA susceptibility. We conducted full-genome characterization and resistance analysis of unusual HCV subtypes in patients receiving DAA treatment. PATIENTS AND METHODS: In this prospective hospital-based study, eligible patients were screened for anti-HCV antibodies and active infection was confirmed by diagnostic 5'NCR-based HCV RNA detection. Genotyping was performed by core region sequencing, and viral load quantified by real-time PCR. For whole genome sequencing, multiplex primers were designed using alignments of global reference sequences. Sequencing was carried out using the Oxford Nanopore Technology platform. Phylogenetic analysis used multiple sequence alignment and the HCV-GLUE resource for resistance-associated substitution (RAS) analysis. RESULTS: Predominant genotype was genotype 3 in 64.3% (n = 45); genotype 6 in 21.4% (n = 15); and genotype 1 in 14.2% (n = 10). Unusual HCV subtype 6xa was detected in two patients and showed no NS5A resistance mutations. One genotype 3b patient relapsed at 24 weeks post-DAA treatment completion and carried NS5A resistance-associated substitutions 30 K and 31 M both at baseline and at relapse, conferring high-level resistance to NS5A inhibitors. CONCLUSION: This is the first report from India of whole genome sequencing of HCV subtype 6xa. The identification of NS5A resistance mutations in the 3b relapse case underscores challenges for global HCV elimination strategies.

Humans

Exploring the mechanism of aroma production in fermented cherry juice by L. brevis LD1.0600 using flavomics and whole genome analysis.

This study focused on L.brevis LD1.0600 with excellent fermentation traits: it analyzed genome-wide key regulatory genes for micro-metabolites, combined with fermented cherry juice flavor metabolomics data, and used machine learning to explore correlations between gene regulation, metabolite production, and flavor formation. The SVM model screened and verified fermented cherry juice VOCs; through OAV and flavor wheel analysis, LD1.0600 emerged as the top-performing strain, with a sweet, fruity dominant aroma. Key aroma-active components (OAV > 100) included 2-methoxy-4-vinylphenol, benzaldehyde, 2-methyl-butanoic acid and hexanoic acid, and 2-methoxy-4-vinylphenol and hexanoic acid elevated by LD1.0600-regulated genes (Chrom1-001884, Chrom1-000925, fabF and Chrom1-000199). At the same time, through research, a "strain screening-SVM screening of DVCs-OAV screening of key aroma components-whole genome sequencing of flavor regulatory genes" system was established. This system can not only be applied to the screen fermentation strains, but also can be extended to the application of other fermentation products.

Fermentation

Complete genome sequence of multidrug-resistant Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from chicken liver, harboring a novel imipenem resistance mechanism.

We present the complete genome sequence of Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from Gallus gallus liver in China, harboring plasmid pSE191. The genome reveals multiple antibiotic resistance mechanisms and phenotypic imipenem resistance without canonical genes.

antibiotic resistance