Search PubMedSearch

SEARCH · Search PubMed

Results for “large genomic fragment”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

798 records · Page 10Linked to original sources

Gloeotrichia echinulata genomes from the United States are nontoxigenic and likely geosmin producers.

Six Gloeotrichia echinulata genomes derived from planktonic harmful algal blooms (HABs) with similar colonial morphology have been sequenced from lakes in the west and northeast regions of USA, four of them to completion. The c. 7 Mbp genomes exhibit a high level of conservation, with 98-99% pairwise genome-wide average nucleotide identity and high levels of synteny, representing a single species cluster. We observed strong conservation of gene clusters responsible for the synthesis of the secondary metabolites and bioactive peptides that are characteristic of HAB-forming cyanobacteria. All six G. echinulata genomes lack genes for the synthesis of classic cyanotoxins, including microcystin, but possess genes responsible for the synthesis of the taste and odor compound geosmin. Interestingly, the geoA geosmin synthase gene in three genomes is homologous to other cyanobacterial geoA genes, while the other three geoA genes are related to actinomyces geoA. Phylogenomic analysis places the G. echinulata genomes within a clade of benthic Nostocales, reflecting an ecological niche featuring extensive growth on the sediment surface before colonies disperse into the epilimnion for planktonic growth. We identify genes conserved in all six genomes that could represent physiological adaptations supporting active growth on sediments and pelagic recruitment independent of wind-driven mixing: phycoerythrin light harvesting complexes for optimal photosynthesis at depth; gliding motility to access patchy nutrient distributions; and gas vesicles with relatively small GvpC proteins that predict resistance to higher hydrostatic pressure. The strong genomic similarity across geographically distant populations suggests that G. echinulata in the United States is a tightly related non-toxigenic species group with predictable properties relevant to public health and drinking water management.

Cyanobacteria

Cytonuclear conflict and reticulate evolution in the Morelloid clade (Solanum, Solanaceae): Insights from genome skimming and network Phylogenomics.

The Morelloid clade (black nightshades) is one of the most strongly supported clades within the megadiverse Solanum genus. It comprises 76 globally distributed, non-spiny herbaceous and suffrutescent species. While often erroneously considered poisonous weeds, several species are economically important as orphan crops. The clade is closely related to tomato and potato but, due to a lack of focused breeding efforts, remains a putative reservoir of genetic diversity for crop improvement. Despite this potential, we lack fundamental knowledge on the evolution of the Morelloid clade. The group includes polyploid species with unknown parental origins-likely reflecting reticulate processes such as hybridization, introgression, and associated backcrossing events. Prior analyses have been unable to disentangle these processes, leaving the mechanisms underlying reticulate evolution in the Morelloid clade poorly understood. Here, we use genome skimming to produce a well-supported maximum likelihood plastid phylogeny from complete circularized plastomes and a coalescent-based species tree from combined Angiosperms353 and conserved ortholog set nuclear markers. Our dataset, composed of previously published data and deep genome skimming from herbarium samples, spans 26 Morelloid species. To investigate phylogenetic discordance, we used a nuclear phylogenetic network, multispecies coalescent simulations, a fused rooted nuclear chloroplast tree, and quantification of nuclear gene tree concordance. We show that incongruence between nuclear and plastid trees is pervasive and cannot be explained by incomplete lineage sorting alone. Instead, our results demonstrate that events consistent with repeated chloroplast capture have shaped the reticulate evolutionary history of the clade, especially among African polyploid and Pan-American diploid lineages.

Phylogeny

Ramu stunt virus genome reveals previously unreported segments and nucleocapsid domain duplication in Mechlorovirus.

Ramu stunt virus (RmSV), a member of the genus Mechlorovirus within the family Phenuiviridae, was previously described as a six-segmented RNA virus infecting sugarcane. In this study, we re-examined type material and additional isolates using high-throughput sequencing and RT-PCR validation, revealing that RmSV possesses a nine-segmented genome, making it the largest reported in the Phenuiviridae. This expanded architecture includes duplicated RNA segments (RNA 2a and RNA 2b) encoding nucleocapsid-like proteins and two novel segments (RNA 7 and RNA 8). Comparative analysis showed that RNA 2a and 2b share about 84% amino acid identity, while RNA 5 encodes a third nucleocapsid homolog, indicating unprecedented domain redundancy. Structural modeling confirmed that all three nucleocapsid proteins maintain a conserved fold despite low sequence identity, with electrostatic mapping suggesting differential RNA-binding potential. Additionally, RNA 6 encodes a hypothetical protein structurally similar to the rice stripe virus disease-specific S-protein, implicating a role in symptom development. Transcript abundance analysis revealed RNA 6 as the most highly expressed segment across isolates. These findings revise the genomic composition of RmSV, highlight mechanisms of genome plasticity and adaptive evolution in plant-infecting bunyaviruses, and underscore practical implications for diagnostic assay design, resistance breeding, and biosecurity surveillance.

Genome, Viral

Genomic and Phenotypic Characterization of Two Novel Enterobacter Phages With EDTA-Enhanced Antibiofilm Activity.

Multidrug-resistant members of the Enterobacter cloacae complex (ECC) are increasingly linked to difficult-to-treat infections and biofilm-mediated antimicrobial tolerance. Here, two lytic phages, vB_EhoIP_HHH and vB_EluM_RZH, displaying podovirus-like and myovirus-like morphology, respectively, were isolated from the River Chelt. HHH has a 39,582 bp genome (51.2% GC, 63 ORFs), while RZH has a 174,197 bp genome (39.4% GC, 314 ORFs), with neither genome carrying antimicrobial resistance, virulence or lysogeny-associated genes. VIRIDIC and VICTOR analyses placed HHH within Kayfunavirus and RZH within Karamvirus, supporting their classification as distinct species. Both phages demonstrated rapid adsorption, short latent periods and stability across physiological pH and temperature ranges. A phage cocktail targeting MDR ECC strain was evaluated with EDTA against established biofilms. Crystal violet assays showed the greatest biomass reduction at MOI 10 with 0.5-0.75 mM EDTA. Bliss independence analysis revealed localized synergy within this window but significant overall antagonism at higher EDTA concentrations. CFU enumeration confirmed greater activity against 24 h than 48 h biofilms. The optimized combination also reduced recoverable bacteria in a fibroblast infection model while maintaining low LDH release. These findings identify two novel lytic Enterobacter phages and support a narrow EDTA concentration window for enhanced phage-mediated antibiofilm activity.

Biofilms

Cost-Effectiveness and the Economics of Genomic Testing and Molecularly Matched Therapies.

Cost-effectiveness analysis of precision oncology can help guide value-driven care. Next-generation sequencing is increasingly cost-efficient over single gene testing because diagnostic algorithms require multiple individual gene tests to determine biomarker status. Matched targeted therapy is often not cost-effective due to the high cost associated with drug treatment. However, genomic profiling can promote cost-effective care by identifying patients who are unlikely to benefit from therapy. Additional applications of genomic profiling such as universal testing for hereditary cancer syndromes and germline testing in patients with cancer may represent cost-effective approaches compared with traditional history-based diagnostic methods.

Humans

Whole genome sequencing of unusual Hepatitis C virus subtypes and drug resistance analysis during direct-acting antiviral therapy in India.

INTRODUCTION AND OBJECTIVES: Pangenotypic direct-acting antivirals (DAA) are effective against highly prevalent Hepatitis C virus (HCV) subtypes, but have been clinically validated almost exclusively in high-income countries. Unusual HCV subtypes may carry natural polymorphisms, potentially impacting DAA susceptibility. We conducted full-genome characterization and resistance analysis of unusual HCV subtypes in patients receiving DAA treatment. PATIENTS AND METHODS: In this prospective hospital-based study, eligible patients were screened for anti-HCV antibodies and active infection was confirmed by diagnostic 5'NCR-based HCV RNA detection. Genotyping was performed by core region sequencing, and viral load quantified by real-time PCR. For whole genome sequencing, multiplex primers were designed using alignments of global reference sequences. Sequencing was carried out using the Oxford Nanopore Technology platform. Phylogenetic analysis used multiple sequence alignment and the HCV-GLUE resource for resistance-associated substitution (RAS) analysis. RESULTS: Predominant genotype was genotype 3 in 64.3% (n = 45); genotype 6 in 21.4% (n = 15); and genotype 1 in 14.2% (n = 10). Unusual HCV subtype 6xa was detected in two patients and showed no NS5A resistance mutations. One genotype 3b patient relapsed at 24 weeks post-DAA treatment completion and carried NS5A resistance-associated substitutions 30 K and 31 M both at baseline and at relapse, conferring high-level resistance to NS5A inhibitors. CONCLUSION: This is the first report from India of whole genome sequencing of HCV subtype 6xa. The identification of NS5A resistance mutations in the 3b relapse case underscores challenges for global HCV elimination strategies.

Humans

The Childhood Cancer and Leukemia International Consortium (CLIC): Expanding global collaboration in pediatric cancer etiology research.

Childhood cancers are rare, but incidence has risen modestly in countries with robust registration, partly reflecting improved diagnosis. In high-income countries, cancer is the leading cause of disease-related death in children. Marked inequities in incidence, survival, and research capacity underscore the need for large-scale collaboration to identify environmental, genetic, and contextual determinants of risk. The Childhood Cancer and Leukemia International Consortium (CLIC) was established in 2007 to study the etiology of childhood leukemia and later expanded in 2019 to include other childhood cancers, principally solid tumors. CLIC pools harmonized, individual-level data from case-control and cohort studies, obtained through interviews, record linkage (insurance claims, registries), or geographic information systems, and integrates germline genomic data where available. Membership has grown from 13 studies in 9 countries to 57 studies in 21 countries; recruitment spans the early 1960s to the present and encompasses approximately 150,000 cases across all tumor types and 300,000 controls with clinical, demographic, and exposure data, centralized via harmonized data dictionaries at the Data Coordination Center, established in 2014 at the International Agency for Research on Cancer, and supported by a secure analysis platform. Pooled analyses across diverse populations have implicated parental age, prenatal vitamin or folic acid use, mode of delivery, fetal growth, selected congenital anomalies, occupational or household exposures (e.g., pesticides), paternal smoking, and markers of early-life immune modulation (e.g., breastfeeding, daycare attendance) in leukemia risk, informing carcinogen evaluation and prevention. The integration of genetic ancestry and germline susceptibility data is clarifying ancestry-related differences in leukemia biology and outcomes, while confirming risk loci with population-specific effects. CLIC is now adding polygenic risk scores and exposomic data to refine etiologic subtyping and identify modifiable pathways, while broadening representation from underserved regions through partnership-building and capacity-strengthening.

Humans

Exploring the mechanism of aroma production in fermented cherry juice by L. brevis LD1.0600 using flavomics and whole genome analysis.

This study focused on L.brevis LD1.0600 with excellent fermentation traits: it analyzed genome-wide key regulatory genes for micro-metabolites, combined with fermented cherry juice flavor metabolomics data, and used machine learning to explore correlations between gene regulation, metabolite production, and flavor formation. The SVM model screened and verified fermented cherry juice VOCs; through OAV and flavor wheel analysis, LD1.0600 emerged as the top-performing strain, with a sweet, fruity dominant aroma. Key aroma-active components (OAV > 100) included 2-methoxy-4-vinylphenol, benzaldehyde, 2-methyl-butanoic acid and hexanoic acid, and 2-methoxy-4-vinylphenol and hexanoic acid elevated by LD1.0600-regulated genes (Chrom1-001884, Chrom1-000925, fabF and Chrom1-000199). At the same time, through research, a "strain screening-SVM screening of DVCs-OAV screening of key aroma components-whole genome sequencing of flavor regulatory genes" system was established. This system can not only be applied to the screen fermentation strains, but also can be extended to the application of other fermentation products.

Fermentation

Complete genome sequence of multidrug-resistant Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from chicken liver, harboring a novel imipenem resistance mechanism.

We present the complete genome sequence of Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from Gallus gallus liver in China, harboring plasmid pSE191. The genome reveals multiple antibiotic resistance mechanisms and phenotypic imipenem resistance without canonical genes.

antibiotic resistance

Introgression shapes the genomic conflict landscape of Malus, providing evidence for a reticulate backbone in a woody crop lineage.

Phylogenomic discordance is widespread across plants, but its evolutionary significance is often obscured when conflict is treated primarily as analytical noise rather than as evidence of underlying processes. In woody lineages in particular, incomplete lineage sorting, introgression, and genome duplication can interact over long timescales to produce complex genomic histories that are not adequately summarized by a strictly bifurcating tree. Here, we use Malus as a model woody genus to investigate how these processes structure conflict across a genus-scale, accession-based phylogenomic framework. Using broad taxon sampling, hundreds of nuclear loci, plastid genomes, and genome-wide SNP summaries, we reconstruct a robust nuclear backbone for sampled Malus lineages and evaluate where discordance is concentrated and which processes best explain it. Nuclear analyses resolve eight major clades, whereas conflict is non-random and localized to recurrent hotspots rather than evenly distributed across the tree. Cytonuclear discordance is similarly concentrated, especially around Clade H, represented by sampled accessions of M. tschonoskii, where localized plastid-nuclear disagreement is consistent with candidate plastid capture or organellar introgression. Multiple complementary analyses further indicate that the strongest conflict is not explained by ILS alone, but instead reflects lineage-structured introgression, while polyploid complexes represent additional localized sources of evolutionary complexity. Together, these results provide evidence for a reticulate genomic backbone in Malus and show how integrating nuclear, plastid, and genome-wide conflict analyses can help distinguish background discordance from process-specific signals in woody plant radiations. Several lineage-level reticulation hypotheses identified here should now be tested with broader population-level sampling and curated reference accessions.

Malus

Shared genetic architecture and cellular convergence between female reproductive disorders and pulmonary function: a genome-wide cross-trait analysis.

Female reproductive disorders (FRDs), including polycystic ovary syndrome, endometriosis, uterine leiomyomata, and infertility, have been epidemiologically associated with impaired pulmonary function. However, it remains unclear whether this cross-organ link reflects shared genetic etiology and, if so, which cellular mechanisms mediate it. We performed a systematic genome-wide cross-trait analysis of three FRDs and lung function traits (FEV₁, FVC, FEV₁/FVC) using GWAS summary statistics from individuals of European ancestry, integrating genetic correlation, bidirectional causal inference, pleiotropy mapping, and single-cell enrichment analyses. We identified significant negative genetic correlations between FRDs and lung volume traits, most prominently for FVC (rg range: - 0.077 to - 0.178). Bidirectional causal analyses indicated that FRDs have a detrimental effect on lung volume, with higher FRD genetic liability associated with reduced lung volume. Cross-trait meta-analysis identified 17 pleiotropic variants across 11 loci, with the 19q13.2 (LTBP4) and 12q13.13 (HOXC6/HOXC9) loci showing strong evidence of shared causal variants. Critically, single-cell analyses revealed that shared genetic risk converged on mesenchymal lineages across organs, specifically alveolar adventitial fibroblasts in the lung and stromal/smooth muscle cells in the endometrium. Transcriptome-wide analyses further nominated the estrogen-responsive gene RERG as a convergent gene linking these conditions with lung function. Our study revealed a shared genetic architecture between female reproductive disorders and lung function traits, providing a basis for further mechanistic investigations and potential clinical evaluation. Furthermore, our findings suggest that shared fibroproliferative and hormone-responsive pathways may offer insights into the biological mechanisms underlying these conditions.

Female

A chromosome-level, haplotype-resolved genome assembly for the barn owl, Tyto alba.

Recent advances in long-read sequencing have enabled near telomere-to-telomere (T2T) assemblies across diverse taxa. However, avian genomes remain challenging due to numerous microchromosomes, small, typically < 20Mb, DNA molecules that are gene-, GC-, and repeat-rich. As a consequence, microchromosomes are often missing from genome assemblies. Here, we present a chromosome-level, haplotype-resolved genome assembly for the Western barn owl (Tyto alba). Using a trio-binning strategy with Illumina parental reads combined with PacBio HiFi and Oxford Nanopore Technologies data, we generated two phased contig sets. These were scaffolded into 40 linkage groups using a linkage map. Comparative analyses identified unplaced HiFi scaffolds corresponding to microchromosomes, which we integrated into six additional microchromosomes using long reads information. The two assemblies present 46 chromosomes, matching the karyotype of the species. They exhibit strong synteny between parental haplotypes, except for a &#x223c;38 Mb complex region on chromosome 7 containing nested inversions. This high-quality reference provides a haplotype-resolved and chromosome-level genome for Strigiformes, enabling fine-scale studies of structural variation and avian genome evolution.

Tyto alba

The Complete Mitochondrial Genome of a Newly Recorded Chinese Species of Diglyphus sabulosus (Hymenoptera: Eulophidae) and Insights into Its Phylogenetic Position.

Diglyphus Walker, 1844 is an economically important genus which many species acting as biocontrol agents against agromyzid leafminer pests, but there is a lack of mitogenomic data on the evolutionary relationships within this genus, hindering a comprehensive understanding of its evolutionary history. We used traditional morphological methods to identify species, and present the first complete mitochondrial genome sequence and characterization of features of Diglyphus sabulosus and further infer its phylogenetic position based on the amino acid sequences of 13 protein-coding genes (PCGs). The complete mitochondrial genome of D. sabulosus is 15,690&#xa0;bp in length, including 13 PCGs, 22 transfer RNA genes, 2 ribosomal RNA genes and a control region. The AT content of the whole genome sequence was 81.0%, indicating a significant AT bias. All protein-coding genes have the typical ATN as the start codon and TAA as the stop codon. Phylogenetic analysis inferred from the amino acid sequences of 13 PCGs revealed that all species within the family Eulophidae constituted a monophyletic clade, supporting the monophyly of this family. D. sabulosus and D. poppoea form a well-supported sister group, representing the species with the closest phylogenetic relationship within the analyzed taxa. In this study, the mitogenome structure was analyzed and the taxonomic status of D. sabulosus was clarified, thus providing a theoretical basis for understanding the phylogenetic relationships of Diglyphus.

Animals