Search PubMedSearch

SEARCH · Search PubMed

Results for “Conserved Sequence”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

403 records · Page 2Linked to original sources

Nanopore-based epigenomic profiling reveals the absence of widespread CpG methylation in the African swine fever virus genome.

DNA methylation is a critical epigenetic mechanism implicated in regulating replication and transcription in DNA viruses. However, the epigenetic landscape of African swine fever virus (ASFV), a large double-stranded DNA virus infecting pigs, remains controversial. Here, we systematically profiled the DNA methylome of the first ASFV strain isolated in Hong Kong (HK_NT_202103) using Oxford Nanopore Technologies (ONT) R10.4.1 sequencing. We employed a paired design: native whole-genome sequencing (WGS) against a methylation-free whole-genome amplification (WGA) control. Using conservative thresholds, we found no evidence of 5-methylcytosine (5mC), especially typical CpG methylation, across the viral genome. Importantly, clear CpG methylation signals were successfully detected in the host genome from WGS data, confirming the functionality of the workflow to detect 5mC at CG sites. While widespread 5mC seems absent, a small number of putative N6-methyladenine (6mA) loci were identified. A specific 6mA candidate exhibited raw ionic current disruptions and gene-level intersection with another ASFV isolate (CAS19-01/2019), although it lacked single-base consensus across different methylation callers or between the two isolates. Although our biological findings are restricted to a single isolate under specific experimental conditions, this study introduces a novel, highly rigorous ONT framework for viral epigenomics research. Furthermore, the absence of ASFV CpG methylation indicates that host CpG-depletion remains a viable strategy for viral metagenomic enrichment. Ultimately, our work offers a critical methodological baseline for ASFV surveillance and highlights the necessity of targeted experimental validation for rare viral modifications.

African Swine Fever Virus

Whole-Genome Deep Learning Predicts Chemotherapy Response in Colorectal Cancer.

Chemotherapy response in colorectal cancer (CRC) exhibits significant heterogeneity, with current clinical predictors failing to capture complex genomic determinants of resistance. We developed a hybrid deep learning framework integrating convolutional neural networks (CNNs) and bidirectional long short-term memory (BiLSTM) networks to analyze whole-genome somatic mutations, evolutionary conservation, chromatin accessibility, and 3D genome architecture in 2,546 TCGA patients. An attention mechanism identified predictive genomic regions. The model achieved an AUC of 0.92 (95% CI: 0.89-0.94) in cross-validation and 0.88 (95% CI: 0.85-0.91) in independent validation, outperforming clinical models (&#x394;AUC = +0.18, p < 0.001). Key predictors included non-coding variants in TP53, KRAS, and PIK3CA regulatory regions. Triple-positive patients (mutations in all 3 regions) had significantly worse progression-free survival (HR = 4.7, p < 0.001). Our framework enables accurate chemotherapy response prediction and reveals novel non-coding resistance mechanisms, advancing precision oncology in CRC.

Humans

Ramu stunt virus genome reveals previously unreported segments and nucleocapsid domain duplication in Mechlorovirus.

Ramu stunt virus (RmSV), a member of the genus Mechlorovirus within the family Phenuiviridae, was previously described as a six-segmented RNA virus infecting sugarcane. In this study, we re-examined type material and additional isolates using high-throughput sequencing and RT-PCR validation, revealing that RmSV possesses a nine-segmented genome, making it the largest reported in the Phenuiviridae. This expanded architecture includes duplicated RNA segments (RNA 2a and RNA 2b) encoding nucleocapsid-like proteins and two novel segments (RNA 7 and RNA 8). Comparative analysis showed that RNA 2a and 2b share about 84% amino acid identity, while RNA 5 encodes a third nucleocapsid homolog, indicating unprecedented domain redundancy. Structural modeling confirmed that all three nucleocapsid proteins maintain a conserved fold despite low sequence identity, with electrostatic mapping suggesting differential RNA-binding potential. Additionally, RNA 6 encodes a hypothetical protein structurally similar to the rice stripe virus disease-specific S-protein, implicating a role in symptom development. Transcript abundance analysis revealed RNA 6 as the most highly expressed segment across isolates. These findings revise the genomic composition of RmSV, highlight mechanisms of genome plasticity and adaptive evolution in plant-infecting bunyaviruses, and underscore practical implications for diagnostic assay design, resistance breeding, and biosecurity surveillance.

Genome, Viral

Genome-wide characterization of heat shock protein genes reveals thermal stress-responsive candidates in Litopenaeus vannamei.

Heat shock proteins (HSPs) are conserved molecular chaperones involved in protein folding, refolding, aggregation prevention, and degradation of damaged proteins. However, the genomic organization and thermal responsiveness of HSP genes in the Pacific white shrimp (Litopenaeus vannamei) remain incompletely understood. Here, we performed a genome-wide analysis of the HSP gene family and examined its phylogenetic relationships, structural features, duplication patterns, sequence variation, interaction networks, and transcriptional responses to acute heat stress. A total of 34 HSP genes were identified and classified into the HSP90, HSP70, HSP40/DNAJ, HSP60, and small HSP families. Phylogenetic, motif, gene structure, synteny, and subcellular localization analyses revealed evolutionary conservation and structural diversification among family members. Three duplicated gene pairs were identified, comprising two segmental duplications and one tandem duplication. All pairs exhibited Ka/Ks ratios below 1, consistent with purifying selection of varying strength. Sequence analysis identified 295 nonsynonymous single-nucleotide polymorphisms, of which 12 were consistently predicted to be deleterious by multiple algorithms. Protein-protein interaction analysis indicated enrichment of protein-folding and cellular stress-response functions. RT-qPCR analysis showed significant induction of HSPA4, HSP90AA1, TRAP1, BiP, and DNAJA1 after 6, 12, and 24&#xa0;h of exposure to 34&#xa0;&#xb0;C, whereas DNAJC3 was significantly induced only at 12&#xa0;h. All six genes reached their highest transcript abundance at 12&#xa0;h. These findings may provide a genomic framework for HSP genes in L. vannamei and identify candidate genes and variants associated with thermal stress responses.

Animals

Meta-PseU: A meta-classifier for robust prediction of RNA pseudouridine modification sites from long sequences.

BACKGROUND AND OBJECTIVES: Pseudouridine (&#x3a8;) represents one of the most abundant and conserved RNA modifications. &#x3a8; provides an additional hydrogen-bond donor that enhances RNA structural stability and modulates translation. It participates in diverse biological processes, including RNA-protein interactions, splicing, translational control, and stress responses. Aberrant pseudouridylation is implicated in cancer, neurodegenerative disorders, and autoimmune diseases. Despite its biological importance, experimental identification of &#x3a8; sites remains time-consuming and costly, limiting the feasibility of transcriptome-wide profiling. Computational approaches have therefore become essential complements to experimental techniques. However, state-of-the-art machine-learning and deep-learning predictors often suffer from limited generalizability due to small training datasets. To overcome these issues, we aim at constructing new long-sequence datasets and developing a novel &#x3a8; site predictor. METHODS: New long-sequence datasets were constructed as benchmarks for RNA &#x3a8;-site prediction. The &#x3a8; modification sites in RMBase 3.0 were mapped to the reference genomes across three species of human, mouse, and yeast, and the RNA sequences with a length of 201 were generated by extending the upstream and downstream from the mapped, central sites. To eliminate sequence redundancy, the sequences were clustered using CD-HIT with a 70% sequence identity threshold. We developed Meta-PseU, a logistic regression-based meta-classifier that considered 118 machine learning and deep learning classifiers. The datasets and programs are freely accessible at https://github.com/kuratahiroyuki/MetaPseU. RESULTS: By optimizing model configuration, we proposed the Meta-PseU model stacking 32 machine learning and deep learning classifiers out of 118 classifiers. Meta-PseU substantially improved model generalizability, overcoming a key limitation of existing approaches. It greatly outperformed state-of-the-art predictors and achieved increasing accuracy with increasing sequence length. CONCLUSIONS: Long-sequence datasets were newly constructed as benchmarks for RNA &#x3a8;-site prediction. Meta-PseU offers a new framework for robust &#x3a8;-site identification by using long sequences.

Pseudouridine

Alternative genetic codes in bacteria and archaea identified with a fast k-mer-based algorithm.

The genetic code is conserved across all domains of life and is often described as universal. Nevertheless, many exceptions to the "universal" code have now been documented, most of these through manual or semiautomated inspection of highly conserved genes. Modern bioinformatics tools improved our ability to find alternative genetic codes but remain computationally expensive, preventing widespread use on thousands of new species identified by sequencing environmental samples. Here, I report a >100-fold accelerated method for inferring the genetic code directly from assembled genomes and apply it to thousands of previously uncharacterized assemblies from archaea and bacteria. I describe three candidate genetic code variations, one of which, an alternative genetic code used by a family of Asgard archaea, is a unique example of sense codon reassignments for this domain. Identifying genetic code variations is important for understanding evolution of the standard code and improving accuracy of protein databases and open reading frame identification.

Genetic Code

Target Capture of Ancient Shell DNA Enables Phylogenetic Reconstruction of Deep-Sea Molluscs.

Target capture is widely used to enrich endogenous DNA from calcium phosphate skeletal material in vertebrates, but its performance on calcium carbonate hard parts widely produced by invertebrates remains poorly understood. Here, we compared DNA recovery from four fresh and 12 ancient (eight radiocarbon-dated to 1671-1135&#x2009;years old before present) deep-sea vesicomyid clam shells, including species Archivesica marissinica, A. nanshaensis and A. okutanii, using whole-genome sequencing (WGS) or target capture of ultraconserved elements (UCEs). WGS achieved 16.65% on-target read recovery of UCEs from fresh soft tissue, but <&#x2009;1% from shell specimens. By contrast, UCE capture in the same specimen increased on-target reads by up to 155-fold, reaching 29.84% in fresh shells and up to 72-fold, reaching 19.89% in ancient shells. Target capture of UCEs recovered 142-1001 loci per sample compared to 0-230 with WGS alone. Ancient shells of A. marissinica and A. okutanii, based on reads mapped with bwa-mem2 and bbmap, exhibited characteristic post-mortem DNA damage signals, with average 5'-end C-to-T misincorporation rates of 3.46% and 15.97%, respectively, exceeding the levels observed in fresh A. marissinica shells (maximum 1.24%). UCE-based phylogenetic reconstructions incorporating shell ancient DNA recovered two major clades within Pliocardiinae, consistent with published phylogenomic trees. Together, these findings demonstrate that target-capture enrichment enables effective recovery of highly degraded DNA from ancient mollusc shells and supports robust phylogenetic inference at the intrageneric scale, expanding the utility of shells-one of the most abundant invertebrate remains-for evolutionary, biogeographic and conservation studies.

Animals

Characteristics and functions of a cell adhesion molecule PvCadN in Penaeus vannamei during WSSV infection.

Cell adhesion not only maintains the integrity of the organism, but also plays an important role in the immune system, which is involved in modulation in the interaction between host and virus. In this study, a novel cell adhesion molecule from Penaeus vannamei, designated as PvCadN, was investigated. It had the typical molecular characteristics of cadherin family, with multiple extracellular cadherin repeat domains, a transmembrane region, and a conserved &#x3b2;-catenin-binding motif. Pvcadn is expressed ubiquitously across all detected tissues, with the highest transcriptional level in gills. RNA interference-mediated silencing of pvcadn significantly impaired the adhesion ability of shrimp hemocytes. Upon WSSV infection, pvcadn showed a tissue-specific expression pattern, with upregulation in gills and downregulation in hemocytes. Knockdown of pvcadn markedly suppressed the transcription of WSSV immediate-early gene ie1 and replication of the viral genome in vivo, suggesting that PvCadN acted as a potential virus-associated molecule. Furthermore, it was found that PvCadN was regulated by Lv&#x3b2;-catenin, a core molecule in the Wnt signaling pathway that functions in innate immunity, at the transcriptional and protein levels. Silencing of lv&#x3b2;-catenin significantly downregulated pvcadn transcription, and Lv&#x3b2;-catenin bound directly to the Cadherin C domain of PvCadN. In summary, the study revealed that PvCadN was a key cell adhesion molecule involved in WSSV infection, which was regulated by Lv&#x3b2;-catenin. Our findings will provide fundamental data for further investigation into cadherin-mediated immune regulation in shrimp, and offer new insights for the prevention and control of WSSV.

Animals

Transcriptomic and RNAi analyses reveal chloride channel 3-associated osmoregulation in Litopenaeus vannamei under low-salinity stress.

Chloride channels and transporters are important for cellular volume regulation and salinity adaptation in euryhaline crustaceans, yet the intestinal transcriptional relationship between plasma-membrane and intracellular chloride pathways remains unclear in Litopenaeus vannamei. In this study, RNA interference of anoctamin 1 (ANO1) was combined with intestinal transcriptome sequencing under the production-relevant low-salinity condition of salinity 3. ANO1 silencing produced a focused transcriptional response, with 16 differentially expressed genes (DEGs) identified (11 upregulated and 5 downregulated). Functional enrichment indicated that these genes were associated with transporter activity, cytoskeletal organization, extracellular matrix-receptor interaction, membrane lipid metabolism, and vesicular processes. Notably, a transcript encoding chloride channel protein 3 (CLC-3) was significantly upregulated following ANO1 knockdown, suggesting a potential transcriptional relationship between ANO1 and CLC-3 in chloride homeostasis. Based on this finding, CLC-3 was selected for full-length cDNA cloning, sequence characterization, salinity-gradient expression analysis, and RNAi-based functional assessment. The cloned CLC-3 cDNA was 2883&#xa0;bp in length and encoded an 850 amino acid protein containing a conserved voltage-gated chloride channel (Voltage-CLC) domain and two cystathionine &#x3b2;-synthase domains. Phylogenetic analysis placed LvCLC-3 within the intracellular CLC-c clade, and tissue distribution analysis showed the highest CLC-3 expression in the intestine. Intestinal CLC-3 expression responded nonlinearly to salinity variation, peaking at salinity 20. Under salinity 3, CLC-3 knockdown reduced ANO1, Na+/K+-ATPase alpha subunit, and Na+-K+-2Cl- cotransporter transcript levels, whereas glutamate-gated chloride channel expression increased. Mild hepatopancreatic structural alterations were also observed after CLC-3 knockdown. These findings suggest that CLC-3 is a salinity-responsive intracellular chloride-transporter candidate associated with intestinal ion-transport-related transcriptional responses after ANO1 suppression in L. vannamei, although the underlying physiological mechanism requires further validation.

Animals

Genome-wide identification and characterization of ABC transporters and their expression in response to saline-alkaline stress and WSSV infection in Fenneropenaeus chinensis.

ATP-binding cassette (ABC) transporters play crucial roles in stress responses across organisms, yet their functions in Fenneropenaeus chinensis remain largely unknown. In this study, we identified 42 FcABC genes (FcABCs) in the F. chinensis genome and analyzed their phylogenetic relationships, gene structures, and chromosomal distributions. Phylogenetic analysis grouped the FcABCs into eight subfamilies (ABCA-ABCH), with conserved motif and domain compositions within each subfamily. Expression analysis showed that several FcABC genes, including FcABCG5, FcABCA1, and FcABCC3, were significantly induced under saline-alkaline stress in gill and hepatopancreas tissues. In contrast, most FcABCs were downregulated after WSSV challenge, though a subset (e.g., FcABCB1, FcABCC1) exhibited early upregulation. Functional validation via RNA interference demonstrated that knockdown of FcABCG5 increased shrimp mortality under saline-alkaline stress. Cis-regulatory element analysis revealed an enrichment of stress- and immune-related elements in FcABC promoters. Protein-protein interaction network predictions indicated potential roles for FcABCs in cholesterol metabolism and organic anion transport. Our findings provide insights into the roles of FcABC genes in stress adaptation and immune defense, offering candidate genes for the breeding of stress-resistant shrimp varieties.

Animals

Functions of tandem-repeat galectins and domain coordination governs galectin-4 activity in grass carp (Ctenopharyngodon idella).

Galectins are &#x3b2;-galactoside-binding lectins that play essential roles in innate immunity. Among them, tandem-repeat galectins (TrGals), typically composed of two distinct carbohydrate-recognition domains (CRDs) connected by a linker peptide, are well established as key regulators of pathogen recognition and host defense in mammals. However, their structural diversity and immunological functions in teleost fish remain poorly understood. In this study, five TrGals (Gal-4, Gal-8a, Gal-8b, Gal-9, and Gal-9like) were identified in grass carp. Sequence and structural analysis revealed that Gal-8a/b, Gal-9, and Gal-9like possess the canonical two-CRD architecture, whereas Gal-4 uniquely contains four highly similar tandem-repeat domains. All five TrGals were broadly expressed across examined tissues, with predominant expression in the liver. Upon Aeromonas hydrophila infection, Gal-4, Gal-8a, Gal-8b, and Gal-9 were rapidly up-regulated at early time points (3-6&#x202f;h). To elucidate the functional significance of CRD number, recombinant full-length CiGal-4 (CiGal4-full) and three truncated variants containing one, two, or three CRDs (CiGal4-1CRD, CiGal4-2CRD, and CiGal4-3CRD) were generated and systematically characterized. All recombinant proteins contained the conserved &#x3b2;-sheet structure typical of galectin CRDs. Functional assays revealed that CiGal4-full displayed the strongest growth-inhibitory activity against all tested bacteria, whereas CiGal4-1CRD showed the weakest effect. Notably, CiGal4-2CRD exhibited the most potent bactericidal activity, surpassing the full-length protein, while CiGal4-3CRD showed no further enhancement. CiGal4-full and CiGal4-2CRD showed superior carbohydrate-binding activities compared with the other variants. Collectively, these results reveal that CRD copy number alone does not linearly determine galectin function. Instead, domain organization and conformational coordination are critical for optimizing antimicrobial activity. This study provides new insights into the structure-function relationships and evolutionary diversification of galectins in teleosts and highlights their potential as novel antimicrobial and immunomodulatory agents in aquaculture.

Animals

Integrating genomic distance analyses in the description of a new family, genus, and species of sponge-associated antipatharians (black corals).

Antipatharians (black corals) are among the least studied coral groups, with much of their diversity still undescribed. Here, we present an integrative morphological, phylogenomic and genomic distance study of deep-sea antipatharians sampled in high seas areas of the North Pacific Ocean and from New Zealand's Exclusive Economic Zone. These corals grow on hexactinellid sponges - a unique characteristic in the order Antipatharia. Using a dataset of ultra-conserved elements and exons, combined with morphological analyses, we reconstruct phylogenomic relationships and formally describe a new family (Eidikopathidae fam. nov.), a new genus (Eidikopathesgen. nov.), and two new species (E. korallispongiasp. nov., E. zealandkoralliasp. nov.). Morphologically, the new family is distinguished by a corallum consisting of a network of loose branches that fuse with the sponge skeletal framework. Phylogenomic analyses recovered consistent topologies with strong nodal support, corroborating the distinct evolutionary placement of this sponge-associated lineage. Pairwise genomic distances estimated using the Tamura-Nei model were concordant with patristic genomic distances, identifying Pteridopathidae as the genetically closest family to Eidikopathidae fam. nov., followed by Myriopathidae and Stylopathidae, which were recovered as sister families in the phylogeny. This pattern shows that genomic distance complements, rather than simply mirrors, tree topology by quantifying accumulated sequence divergence among lineages. Together, these results provide the first genomic distance framework for Antipatharia, offering a baseline for future systematic, evolutionary, and biodiversity studies on this fundamental shallow, mesophotic and deep-sea coral group.

Animals

Gloeotrichia echinulata genomes from the United States are nontoxigenic and likely geosmin producers.

Six Gloeotrichia echinulata genomes derived from planktonic harmful algal blooms (HABs) with similar colonial morphology have been sequenced from lakes in the west and northeast regions of USA, four of them to completion. The c. 7 Mbp genomes exhibit a high level of conservation, with 98-99% pairwise genome-wide average nucleotide identity and high levels of synteny, representing a single species cluster. We observed strong conservation of gene clusters responsible for the synthesis of the secondary metabolites and bioactive peptides that are characteristic of HAB-forming cyanobacteria. All six G. echinulata genomes lack genes for the synthesis of classic cyanotoxins, including microcystin, but possess genes responsible for the synthesis of the taste and odor compound geosmin. Interestingly, the geoA geosmin synthase gene in three genomes is homologous to other cyanobacterial geoA genes, while the other three geoA genes are related to actinomyces geoA. Phylogenomic analysis places the G. echinulata genomes within a clade of benthic Nostocales, reflecting an ecological niche featuring extensive growth on the sediment surface before colonies disperse into the epilimnion for planktonic growth. We identify genes conserved in all six genomes that could represent physiological adaptations supporting active growth on sediments and pelagic recruitment independent of wind-driven mixing: phycoerythrin light harvesting complexes for optimal photosynthesis at depth; gliding motility to access patchy nutrient distributions; and gas vesicles with relatively small GvpC proteins that predict resistance to higher hydrostatic pressure. The strong genomic similarity across geographically distant populations suggests that G. echinulata in the United States is a tightly related non-toxigenic species group with predictable properties relevant to public health and drinking water management.

Cyanobacteria

Suppression of HIV-1 replication in CEM-A cell cultures by trans-splicing group I introns targeting PAS/PBS sequences and conditionally expressing &#x394;N-Bax.

Anti-HIV group I introns containing antisense guide sequences directed against the HIV-1 primer activation signal and primer-binding site (PAS/PBS) were designed and evaluated. Because PAS/PBS sequences are present in the viral RNA species examined, these RNAs can serve as trans-splicing substrates. The introns were active against both artificial target RNAs and viral RNA generated during infection. Cleavage and degradation of targeted viral RNA may have contributed to suppression, whereas inclusion of a 3' exon encoding the proapoptotic protein &#x394;N-Bax was associated with increased programmed cell death and may have augmented suppression of viral replication. In cultured CEM-A cells, transgene expression of these introns markedly suppressed HIV-1 replication, with p24 levels falling below the assay detection limit in selected clones. RESULTS: RT-PCR and sequence analysis detected splice products containing the expected PAS/PBS junctions. In the dual-luciferase assay, intron expression reduced normalized Gaussia luciferase signal by approximately 70% relative to the negative control. Qualitative Annexin V imaging and caspase-3 assays were consistent with infection-dependent apoptosis after &#x394;N-Bax splice-product formation. Transient expression of each intron in HEK293T cells followed by infection with VSV-G-pseudotyped HIV-1NL4-3&#x202f;at an MOI of 2 reduced p24 levels by approximately 50% at 4 days post-infection. Construct 128L produced the strongest RT-PCR band under the tested conditions and was selected for subsequent experiments. A canonical splice product and a low-abundance noncanonical splice product were detected; both involved the intended HIV-derived target RNA, although transcriptome-wide off-target splicing was not assessed. Heterogeneous transformed HEK293T populations showed an approximately 2-log10 reduction in p24. In selected clonal HEK293T and CEM-A lines, p24 was below the assay detection limit at the measured endpoints, including up to 90 days after infection in some CEM-A clones. CONCLUSIONS: PAS/PBS-targeting group I introns suppressed HIV-1-associated p24 production in the tested cell-culture models. Linking the introns to a &#x394;N-Bax 3' exon was associated with infection-dependent apoptosis and may further limit viral replication and spread. The use of highly conserved, functionally constrained target sequences may reduce the likelihood of escape, but viral evolution and transcriptome-wide off-target effects were not assessed. This conditional death-upon-infection strategy warrants further evaluation in primary-cell and in vivo models.

Humans

Cytonuclear conflict and reticulate evolution in the Morelloid clade (Solanum, Solanaceae): Insights from genome skimming and network Phylogenomics.

The Morelloid clade (black nightshades) is one of the most strongly supported clades within the megadiverse Solanum genus. It comprises 76 globally distributed, non-spiny herbaceous and suffrutescent species. While often erroneously considered poisonous weeds, several species are economically important as orphan crops. The clade is closely related to tomato and potato but, due to a lack of focused breeding efforts, remains a putative reservoir of genetic diversity for crop improvement. Despite this potential, we lack fundamental knowledge on the evolution of the Morelloid clade. The group includes polyploid species with unknown parental origins-likely reflecting reticulate processes such as hybridization, introgression, and associated backcrossing events. Prior analyses have been unable to disentangle these processes, leaving the mechanisms underlying reticulate evolution in the Morelloid clade poorly understood. Here, we use genome skimming to produce a well-supported maximum likelihood plastid phylogeny from complete circularized plastomes and a coalescent-based species tree from combined Angiosperms353 and conserved ortholog set nuclear markers. Our dataset, composed of previously published data and deep genome skimming from herbarium samples, spans 26 Morelloid species. To investigate phylogenetic discordance, we used a nuclear phylogenetic network, multispecies coalescent simulations, a fused rooted nuclear chloroplast tree, and quantification of nuclear gene tree concordance. We show that incongruence between nuclear and plastid trees is pervasive and cannot be explained by incomplete lineage sorting alone. Instead, our results demonstrate that events consistent with repeated chloroplast capture have shaped the reticulate evolutionary history of the clade, especially among African polyploid and Pan-American diploid lineages.

Phylogeny

Dissemination of blaKPC-3-harbouring Klebsiella pneumoniae across ST48 and ST628 in multiple healthcare facilities in the Republic of Korea.

Klebsiella pneumoniae carbapenemase-3 (KPC-3) remains rare in South Korea, where KPC-2 is the dominant carbapenemase, making the repeated detection of a concentrated blaKPC-3 signal over five years notable. We performed genomic analyses of blaKPC-3-harbouring K. pneumoniae from a regional healthcare network. Two chromosomally distinct lineages with concordant capsule loci (ST628/KL15 and ST48/KL62) presented multidrug-resistant phenotypes, and the virulence-associated loci were confined to ST48. Single-nucleotide polymorphism (SNP) analyses revealed near-clonal relatedness within lineages, with 0-38 pairwise SNPs among ST628 isolates and 8 SNPs between the two ST48 isolates. Core-genome multilocus sequence typing (cgMLST) supported this structure, as ST628 isolates were assigned to complex type 19149 with 0-7 allelic differences, and ST48 isolates were assigned to complex type 19150 with 5 allelic differences. These patterns support vertical spread via clonal expansion across multiple facilities. Despite substantial chromosomal separation, most isolates carried the same IncFII(K) plasmid backbone and blaKPC-3, and they were nearly indistinguishable from a plasmid previously reported in South Korea. One isolate carried blaKPC-3 on a distinct multireplicon IncFIB(K)/IncFII(K) plasmid, indicating that the signal was not confined to a single plasmid backbone. In both plasmids, blaKPC-3 was embedded within Tn4401b. These findings indicate that a rare blaKPC-3 genotype can persist regionally through sustained clonal dissemination and that cross-lineage linkage is compatible with past horizontal transfer involving a conserved plasmid. These findings underscore the need for subtype-resolved, regionally coordinated genomic surveillance in connected healthcare networks to detect uncommon carbapenemase variants early.

Klebsiella pneumoniae

Diversity and population connectivity of members of the family Eunicidae inhabiting deep-water corals in the North Atlantic.

Eunicid polychaetes are often found in association with Cold Water Corals (CWCs), even establishing symbiotic relationships, such as those described between Desmophyllum pertusum and Eunice norvegica. While genetic connectivity of CWCs across the North Atlantic has been widely studied, little is known about their associated fauna in this regard. Here, we present a study combining a focused analysis of the genetic and genomic connectivity of E. norvegica with a regional assessment of the distribution and evolutionary relationships of three CWC-associated eunicid species from the Cantabrian Sea and the North of the United Kingdom (190-1,230&#xa0;m depth). An integrative approach using genetic (16S, COI and 18S), morphological and ecological data allowed the identification of the eunicids studied, with new records of Eunice cf. nicidioformis and Leodice cf. antarctica in the Cantabrian Sea, as well as previously undocumented associations with CWC species. In addition, RADseq data contributed to the delimitation of the closely related species E. norvegica and Eunice philocorallia. Moreover, the genetic connectivity of E. norvegica was studied trough a RADseq (1,067 neutral SNPs) approach. Our results indicate a single panmictic population across approximately 2,000&#xa0;km, suggesting that oceanographic currents facilitate passive dispersal of E. norvegica lecithotrophic larvae, aided by coral host stepping-stones. The connectivity patterns observed for E. norvegica mirror those of D. pertusum, on which the worm is ecologically dependent. Our study highlights the importance of using integrated genetic, morphological and ecological data to characterise and delineate understudied CWC-associated species and improve our understanding of their dispersal capabilities and genetic connectivity to inform future conservation recommendations.

Animals

Phenotypic and transcriptomic characterization of biallelic RNU2-2 developmental and epileptic encephalopathy.

OBJECTIVE: A significant proportion of individuals with suspected genetic developmental and epileptic encephalopathies (DEEs) remain unsolved following whole genome sequencing (WGS). Here we describe biallelic RNU2-2 variants causing a recently reported, severe, recessive DEE. METHODS: We screened individuals who have received WGS analyses at the Genomic Medicine Centre Karolinska for Rare Diseases for biallelic RNU2-2 variants. Deep phenotyping was performed through reviewing entire medical histories and phenotypic traits were transcribed to their corresponding Human Phenotype Ontology (HPO) term. HPO terms were used to generate pairwise phenotypic similarity scores and assess for significantly shared phenotype enrichment in the RNU2-2 sub-cohort. RNA sequencing analyses were performed in fibroblast and blood tissues to compare splicing events between RNU2-2 individuals and two independent control groups. RESULTS: We identified 14 individuals from nine families with 12 ultra-rare biallelic RNU2-2 variants clustering in the conserved 5' domains. Genotype data from 13 of 14 individuals has been reported previously as part of a larger cohort. All individuals presented with a highly concordant, severe DEE, characterized by severe to profound intellectual disability, inability to walk or communicate, hyperkinesia, and refractory seizures. Infantile spasms and tonic seizures were the predominant seizure types and a Lennox-Gastaut syndrome-like phenotype was common. These individuals had a significantly similar phenotypic signature when compared with 703 individuals with complex pediatric epilepsies (two-sided Monte Carlo permutation test, p&#x2009;=&#x2009;.005). RNA sequencing analyses showed aberrant splicing, with the most pronounced effects in fibroblast tissues in mutually exclusive exon and alternate 3' splice-site events, which were not detectable in blood. SIGNIFICANCE: We present deep phenotyping data and transcriptomic analyses that provide support for rare, 5' clustering biallelic RNU2-2 variants causing this novel, severe DEE. We propose an RNA sequencing methodology on fibroblast tissue for future validation of RNU2-2 variants.

autosomal recessive disease