Search PubMedSearch

SEARCH · Search PubMed

Results for “Molecular Sequence Data”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,346 records · Page 8Linked to original sources

Recovery of polysaccharides from marc and pomace through sequential extractions assisted by ultrasound, enzymes and acid maceration.

This study evaluated the pilot-scale recovery of polysaccharides from Vitis vinifera pomace/marc using sequential extraction strategies combining high-power ultrasound (UAE), enzymes (EAE), and acid maceration (AAE). Laboratory-scale trials identified optimal conditions for enzyme dosage and liquid/solid ratio (L/S). Pilot-scale trials demonstrated that the extraction sequence and the processing byproducts influenced extraction efficiency, total soluble polysaccharide in the extract (TSP), and polysaccharide composition. Post-maceration at pH 3.2, with/without the maximum enzyme dose after UAE in a L/S of 1.3/1, improved structural polysaccharide extraction from Viura pomace, while Tempranillo marc showed better recovery of pectic families and TSP with UAE + EAE. Separating grape pomace extract (UAE) from the post-maceration stage at pH 3.2 produced two extracts: E1, with higher yield (19.9%), enriched in structural polysaccharides and oligosaccharides, and E2, enriched in high and medium molecular weight pectic polysaccharides (58.03%), a low degree of esterification (17.1%) and more complex rhamnogalacturan structures.

Polysaccharides

Phenotypic and transcriptomic characterization of biallelic RNU2-2 developmental and epileptic encephalopathy.

OBJECTIVE: A significant proportion of individuals with suspected genetic developmental and epileptic encephalopathies (DEEs) remain unsolved following whole genome sequencing (WGS). Here we describe biallelic RNU2-2 variants causing a recently reported, severe, recessive DEE. METHODS: We screened individuals who have received WGS analyses at the Genomic Medicine Centre Karolinska for Rare Diseases for biallelic RNU2-2 variants. Deep phenotyping was performed through reviewing entire medical histories and phenotypic traits were transcribed to their corresponding Human Phenotype Ontology (HPO) term. HPO terms were used to generate pairwise phenotypic similarity scores and assess for significantly shared phenotype enrichment in the RNU2-2 sub-cohort. RNA sequencing analyses were performed in fibroblast and blood tissues to compare splicing events between RNU2-2 individuals and two independent control groups. RESULTS: We identified 14 individuals from nine families with 12 ultra-rare biallelic RNU2-2 variants clustering in the conserved 5' domains. Genotype data from 13 of 14 individuals has been reported previously as part of a larger cohort. All individuals presented with a highly concordant, severe DEE, characterized by severe to profound intellectual disability, inability to walk or communicate, hyperkinesia, and refractory seizures. Infantile spasms and tonic seizures were the predominant seizure types and a Lennox-Gastaut syndrome-like phenotype was common. These individuals had a significantly similar phenotypic signature when compared with 703 individuals with complex pediatric epilepsies (two-sided Monte Carlo permutation test, p = .005). RNA sequencing analyses showed aberrant splicing, with the most pronounced effects in fibroblast tissues in mutually exclusive exon and alternate 3' splice-site events, which were not detectable in blood. SIGNIFICANCE: We present deep phenotyping data and transcriptomic analyses that provide support for rare, 5' clustering biallelic RNU2-2 variants causing this novel, severe DEE. We propose an RNA sequencing methodology on fibroblast tissue for future validation of RNU2-2 variants.

autosomal recessive disease

Genetic diversity and recombination of NA-PRRSV field strains in Vietnam: Implications for vaccine efficacy.

Porcine reproductive and respiratory syndrome (PRRS) causes severe reproductive losses in pregnant sows and piglets, resulting in substantial economic impact on the swine industry worldwide. However, due to the significant genetic diversity and rapid evolutionary changes of the pathogen, continuous surveillance and detailed genetic analysis of circulating strains are essential. The current study aimed to evaluate the genetic diversity of the hypervariable (HV) region of non-structural protein 2 (nsp2) among North American PRRSV strains isolated from swine farms in Vietnam. Phylogenetic analysis and multiple sequence alignment were conducted to determine subtype classification and assess genetic variability. A total of 48 field isolates were obtained, of which 12.5% belonged to classical NA-PRRSV, 16.6% to NADC30-like and 70.9% to HP-PRRSV, primarily distributed across sublineages 1.4, 5.1, 8.7 and 8.9. Amino acid comparisons found multiple insertions, deletions and substitutions at various positions within the hypervariable region of nsp2. The study revealed substantial genetic variation in the HV region of nsp2 among NA-PRRSV field strains, largely associated with recombination and immune escape. These findings highlight epidemiological risks to vaccine efficacy and underscore the need for continuous molecular surveillance to support effective PRRSV control in Vietnam.

PRRSV

The Complete Mitochondrial Genome of a Newly Recorded Chinese Species of Diglyphus sabulosus (Hymenoptera: Eulophidae) and Insights into Its Phylogenetic Position.

Diglyphus Walker, 1844 is an economically important genus which many species acting as biocontrol agents against agromyzid leafminer pests, but there is a lack of mitogenomic data on the evolutionary relationships within this genus, hindering a comprehensive understanding of its evolutionary history. We used traditional morphological methods to identify species, and present the first complete mitochondrial genome sequence and characterization of features of Diglyphus sabulosus and further infer its phylogenetic position based on the amino acid sequences of 13 protein-coding genes (PCGs). The complete mitochondrial genome of D. sabulosus is 15,690 bp in length, including 13 PCGs, 22 transfer RNA genes, 2 ribosomal RNA genes and a control region. The AT content of the whole genome sequence was 81.0%, indicating a significant AT bias. All protein-coding genes have the typical ATN as the start codon and TAA as the stop codon. Phylogenetic analysis inferred from the amino acid sequences of 13 PCGs revealed that all species within the family Eulophidae constituted a monophyletic clade, supporting the monophyly of this family. D. sabulosus and D. poppoea form a well-supported sister group, representing the species with the closest phylogenetic relationship within the analyzed taxa. In this study, the mitogenome structure was analyzed and the taxonomic status of D. sabulosus was clarified, thus providing a theoretical basis for understanding the phylogenetic relationships of Diglyphus.

Animals

Getting to the Core of the Matter-Assessing the Role of Replication in Metabarcoding-Based sedaDNA.

Replication is central to most experimental and sampling designs, increasing inferential power and capturing fine-scale data heterogeneity. However, its importance remains poorly evaluated in some ecological and evolutionary settings. This is the case of metabarcoding studies using DNA recovered from sedimentary archives, in which biological signals integrate ecological information through depositional and burial processes, yet are commonly inferred from a single sediment core per site. Here, we evaluated the effect of different types of replication using sedimentary DNA metabarcoding data from two genetic markers (mitochondrial COI and nuclear 18S) using a nested sampling design. The design included three intertidal sites, three spatially separated sediment cores per site (biological replicates), two sediment horizons per core, and eight PCR (technical) replicates per sediment sample. Variance partitioning showed that site identity and sediment age group together explained > 70% of the variation in beta diversity, indicating that among-site spatial and stratigraphic differences were the dominant drivers of community composition. PERMANOVA likewise identified non-significant effects of biological replication. Among PCR replicates from the same sediment sample, richness varied substantially, whereas Shannon diversity was more consistent. Despite this variability, differences in community composition among technical replicates remained smaller than those associated with biological replication or site identity, indicating a limited influence on broader ecological patterns. Community composition was highly similar among replicate cores within sites, consistent with stratigraphic coherence. These results indicate limited within-site heterogeneity and suggest that, under stratigraphically coherent conditions, increasing biological replication may provide little additional information, whereas enhancing technical replication and stratigraphic resolution can improve ecological inference from sedimentary DNA metabarcoding datasets.

DNA Barcoding, Taxonomic

Complete genome sequence of the Anaplasma phagocytophilum clinical isolate NCH-1.

Anaplasma phagocytophilum is an obligate intracellular gram-negative bacterium and etiologic agent of human granulocytic anaplasmosis. A. phagocytophilum genomic sequencing has historically been performed via short-read platforms. Our optimized bacterial isolation protocol combined with Nanopore sequencing produced a single, closed 1,481,805 bp circular A. phagocytophilum strain NCH-1 chromosome.

Anaplasma phagocytophilum

[Analysis of clinical phenotypes and pathogenicity of a c.4476+5G>T variant of SCN1A gene in a Chinese pedigree affected with Genetic epilepsy with febrile seizures plus].

OBJECTIVE: To explore the pathogenicity and characteristics of a heterozygous splicing variant of SCN1A gene in a Chinese pedigree affected with Genetic epilepsy with febrile seizures plus (GEFS+). METHODS: A retrospective analysis was carried out on the clinical data and results of genetic testing of a GEFS+ pedigree consisting of 5 members who had visited the First Affiliated Hospital of Zhengzhou University on July 1, 2024. Pathogenicity of the splicing variant of the SCN1A gene was validated with a minigene splicing assay. This study was approved by the Medical Ethics Committee of the the First Affiliated Hospital of Zhengzhou University (Ethics No.: KS-2018-KY-36). RESULTS: The proband, a 24-year-old female, presented with FS in conjunct with focal seizures, and both of her younger brothers had Dravet syndrome. All of the three patients had carried a c.4476+5G>T variant of the SCN1A gene, which was unreported previously. Minigene experiment verified that the variant could cause loss of the first 7 bps of exon 24 and 138 bps from exon 23 of the SCN1A gene, resulting in alteration p.V1447_1495delfs*6 and affecting splicing. Based on the guidelines from American College of Medical Genetics and Genomics (ACMG), the variant was predicted as likely pathogenic (PVS1+PM2_Supporting). CONCLUSION: The c.4476+5G>T variant at an intronic site of the SCN1A gene probably underlay the pathogenesis of GEFS+ in this pedigree.

Adult

Draft genome sequence of Enterococcus casseliflavus strain MBBL_MP4 isolated from healthy bovine milk.

We report the draft genome sequence of Enterococcus casseliflavus MBBL_MP4, recovered from healthy bovine milk. The 3.45-Mbp genome assembly comprises 27 contigs and indicates low pathogenic potential, with no acquired antimicrobial resistance or known virulence genes. This genome provides a valuable resource for the genomic characterization of bovine-associated E. casseliflavus.

Enterococcus casseliflavus

Genomic characterization of a hypervirulent Aeromonas veronii NN0115 from Nile tilapia and head kidney transcriptome of infected fish reveals B-cell-dominated immune response with specific immunoglobulin downregulation.

Aeromonas veronii is a pathogen of multiple fish species, yet systematic understanding of its infection in Nile tilapia (Oreochromis niloticus) remains limited. A dominant strain, NN0115, was isolated from a natural outbreak and identified as A. veronii by 16S rRNA and whole-genome average nucleotide identity (ANI, 96.33%). Experimental infection revealed high virulence (LD50 = 3.41 × 106 CFU/mL, equivalent to 8.53 × 104 CFU/fish). The genome is 4.58 Mb (58.57% GC) and encodes 4216 proteins. Virulence factor analysis identified 1253 genes, dominated by motility-related (264) and immune modulation (208) factors. Genomic island GI2 harbors 7 virulence genes and two dual-function resistance-virulence genes. The strain is resistant to 9 of 25 agents tested but carries three RND efflux pump genes whose predicted resistance was not phenotypically observed. The head kidney transcriptome of tilapia at 24 h post-bacterial infection identified 773 differentially expressed genes; among them, 57 were immunoglobulin (Ig) genes, and 56 were down-regulated. Integration of published single-cell transcriptomic data showed that non-Ig B-cell marker genes were down-regulated by 32%, whereas Ig genes were reduced by 63%, indicating selective transcriptional suppression of Ig genes rather than a general decrease in B-cell transcriptional activity. Together, this study provides a comprehensive characterization of a highly virulent A. veronii from Nile tilapia and reveals that selective downregulation of B-cell Ig genes is the dominant transcriptional feature of the host head kidney response.

Animals

Diagnostic and clinical utility of exome sequencing and chromosomal microarray in children with GDD/iD: a meta-analysis.

BACKGROUND: Global developmental delay/intellectual disability (GDD/ID) is among the most common neurodevelopmental disorders, with up to half of cases are attributed to genetic factors. Chromosome microarray (CMA) has traditionally been the primary genetic test for idiopathic GDD/ID. However, whole exome sequencing (WES) and whole genome sequencing (WGS) have recently emerged, substantially increasing diagnostic yields in these populations. METHODS: We conducted a comprehensive literature search of PubMed, Scopus, EMBASE, and the Cochrane Library from inception to April 29, 2025. Studies reporting the diagnostic utility of these tests in children with GDD/ID were included and analyzed. RESULTS: A total of 102 studies, comprising 55,752 children, were reviewed. The pooled diagnostic yield of WES was 0.37 (95% CI: 0.33-0.41; I2 = 93%), significantly higher than that of CMA at 0.19 (95% CI: 0.16-0.21; I2 = 95%). Subgroup analyses showed that WES yielded significantly higher diagnostic rates than CMA in both same-sample comparisons (OR = 2.27, 95% CI: 1.08-4.78) and different-sample comparisons (OR = 1.65, 95% CI: 1.15-2.37). Only one study evaluated WGS, reporting a diagnostic yield of 0.27. Meta-regression revealed a significant association between CMA diagnostic yield and the proportion of male participants (p&#x2009;<&#x2009;0.01), but not with WES. No significant difference in diagnostic utility was observed between isolated GDD/ID and GDD/ID with comorbidities. CONCLUSION: In children with unexplained GDD/ID, WES demonstrates superior diagnostic and clinical utility compared to CMA. Incorporating WES as a first-line investigation in the diagnostic evaluation of GDD/ID may be warranted.

Humans

Multi-omics analysis reveals coordinated epigenetic dysregulation in atrazine-induced dopaminergic neurotoxicity.

Atrazine (ATR), a widely used triazine herbicide, has been linked to neurotoxicity, yet the epigenetic mechanisms underlying its dopaminergic effects remain unclear. This study investigated whether coordinated miRNA dysregulation and DNA methylation alterations contribute to ATR-induced Parkinson's disease (PD)-like neurotoxicity. Male Sprague-Dawley rats were administered ATR (50&#x202f;mg/kg/day) for 90 days, resulting in motor and cognitive deficits with dopaminergic dysfunction, including increased &#x3b1;-synuclein and reduced tyrosine hydroxylase expression. Small RNA sequencing identified 72 differentially expressed miRNAs in the substantia nigra, enriched in PI3K-Akt, MAPK, and Ras signaling pathways. In a cohort of six PD patients and six matched controls, genome-wide DNA methylation profiling revealed 4694 differentially methylated positions, predominantly hypomethylated, with overlapping enrichment in neuronal signaling pathways. Weighted gene co-expression network analysis identified a PD-associated module strongly correlated with disease status (r&#x202f;=&#x202f;-0.95, P&#x202f;<&#x202f;0.001). Multi-omics integration identified CASP3 as a central hub gene. External validation supported CASP3 relevance in PD (AUC&#x202f;=&#x202f;0.833), and molecular docking suggested potential ATR-CASP3 interaction. Further analysis predicted upregulated miR-3552 as a potential upstream regulator of CASP3. These findings indicate that ATR-induced neurotoxicity may be mediated through the miR-3552/CASP3 signaling axis, ultimately regulating apoptosis and contributing to neurodegeneration.

Animals

Whole genome sequencing of unusual Hepatitis C virus subtypes and drug resistance analysis during direct-acting antiviral therapy in India.

INTRODUCTION AND OBJECTIVES: Pangenotypic direct-acting antivirals (DAA) are effective against highly prevalent Hepatitis C virus (HCV) subtypes, but have been clinically validated almost exclusively in high-income countries. Unusual HCV subtypes may carry natural polymorphisms, potentially impacting DAA susceptibility. We conducted full-genome characterization and resistance analysis of unusual HCV subtypes in patients receiving DAA treatment. PATIENTS AND METHODS: In this prospective hospital-based study, eligible patients were screened for anti-HCV antibodies and active infection was confirmed by diagnostic 5'NCR-based HCV RNA detection. Genotyping was performed by core region sequencing, and viral load quantified by real-time PCR. For whole genome sequencing, multiplex primers were designed using alignments of global reference sequences. Sequencing was carried out using the Oxford Nanopore Technology platform. Phylogenetic analysis used multiple sequence alignment and the HCV-GLUE resource for resistance-associated substitution (RAS) analysis. RESULTS: Predominant genotype was genotype 3 in 64.3% (n = 45); genotype 6 in 21.4% (n = 15); and genotype 1 in 14.2% (n = 10). Unusual HCV subtype 6xa was detected in two patients and showed no NS5A resistance mutations. One genotype 3b patient relapsed at 24 weeks post-DAA treatment completion and carried NS5A resistance-associated substitutions 30 K and 31 M both at baseline and at relapse, conferring high-level resistance to NS5A inhibitors. CONCLUSION: This is the first report from India of whole genome sequencing of HCV subtype 6xa. The identification of NS5A resistance mutations in the 3b relapse case underscores challenges for global HCV elimination strategies.

Humans

The cold case of state transition 7 (stt7) mutants of Chlamydomonas reinhardtii, solved by whole-genome sequencing.

The process of State Transitions (ST) corresponds to an STT7 kinase-driven redistribution of the transmembrane LHCII antenna proteins between Photosystem II (PSII) and Photosystem I (PSI), which results from changes in their phosphorylation state. For the past two decades, two LHCII-kinase mutants, stt7-1 and stt7-9, have been instrumental in the study of STs in Chlamydomonas reinhardtii, the former being a null mutant for the kinase but quasi-sterile in crosses, while the latter, although fertile, has a leaky phenotype. Using long-read sequencing, this study further characterized the genetic lesions of the stt7 mutant strains through whole-genome reconstruction and de novo chromosome assembly. In addition, two new stt7 null mutants were generated, one derived by crosses from the original stt7-1 and one obtained by Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR)-associated protein 9 (Cas9) technology. This work provides a comprehensive genomic characterization of the original stt7-1 null mutant, revealing extensive chromosomal rearrangements and high levels of aneuploidy, associated with increased cell size and meiotic dysfunction. Reassessment of their physiology and genetic backgrounds highlights the need for caution in interpreting genetic information. We thus produced more reliable null mutants for the LHCII-kinase, amenable to genetic crosses for the study of STs in a variety of genetic backgrounds.

Chlamydomonas reinhardtii

Complete genome sequence of multidrug-resistant Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from chicken liver, harboring a novel imipenem resistance mechanism.

We present the complete genome sequence of Salmonella enterica subsp. enterica serovar Enteritidis SD191 isolated from Gallus gallus liver in China, harboring plasmid pSE191. The genome reveals multiple antibiotic resistance mechanisms and phenotypic imipenem resistance without canonical genes.

antibiotic resistance

Meta-PseU: A meta-classifier for robust prediction of RNA pseudouridine modification sites from long sequences.

BACKGROUND AND OBJECTIVES: Pseudouridine (&#x3a8;) represents one of the most abundant and conserved RNA modifications. &#x3a8; provides an additional hydrogen-bond donor that enhances RNA structural stability and modulates translation. It participates in diverse biological processes, including RNA-protein interactions, splicing, translational control, and stress responses. Aberrant pseudouridylation is implicated in cancer, neurodegenerative disorders, and autoimmune diseases. Despite its biological importance, experimental identification of &#x3a8; sites remains time-consuming and costly, limiting the feasibility of transcriptome-wide profiling. Computational approaches have therefore become essential complements to experimental techniques. However, state-of-the-art machine-learning and deep-learning predictors often suffer from limited generalizability due to small training datasets. To overcome these issues, we aim at constructing new long-sequence datasets and developing a novel &#x3a8; site predictor. METHODS: New long-sequence datasets were constructed as benchmarks for RNA &#x3a8;-site prediction. The &#x3a8; modification sites in RMBase 3.0 were mapped to the reference genomes across three species of human, mouse, and yeast, and the RNA sequences with a length of 201 were generated by extending the upstream and downstream from the mapped, central sites. To eliminate sequence redundancy, the sequences were clustered using CD-HIT with a 70% sequence identity threshold. We developed Meta-PseU, a logistic regression-based meta-classifier that considered 118 machine learning and deep learning classifiers. The datasets and programs are freely accessible at https://github.com/kuratahiroyuki/MetaPseU. RESULTS: By optimizing model configuration, we proposed the Meta-PseU model stacking 32 machine learning and deep learning classifiers out of 118 classifiers. Meta-PseU substantially improved model generalizability, overcoming a key limitation of existing approaches. It greatly outperformed state-of-the-art predictors and achieved increasing accuracy with increasing sequence length. CONCLUSIONS: Long-sequence datasets were newly constructed as benchmarks for RNA &#x3a8;-site prediction. Meta-PseU offers a new framework for robust &#x3a8;-site identification by using long sequences.

Pseudouridine

Evaluation of one-step amplicon-based targeted enrichment for SARS-CoV-2 whole-genome sequencing using the Midnight amplicon scheme.

Genomic surveillance proved invaluable during the COVID-19 pandemic for tracking SARS-CoV-2 variants and guiding outbreak responses, underscoring the ongoing need to reduce whole-genome sequencing (WGS) costs and improve workflow efficiency to ensure accessibility in resource limited settings. Here, we evaluated a one-step reverse transcription polymerase chain reaction (RT-PCR) approach using the Midnight V2 primer scheme for targeted amplification of the SARS-CoV-2 genome, assessed its compatibility with Illumina sequencing, and compared its performance to a well-established two-step method. Initially, we determined optimal RT-PCR reaction conditions using the Midnight V2 primer panel for the one-step RT-PCR kit and scaled reaction volumes for both RT-PCR and library preparation. Clinical specimens (n&#x202f;=&#x202f;53) that had undergone routine WGS for surveillance purposes using the established two-step RT-PCR method were compared using the one-step RT-PCR assay. For samples with genome completeness greater than 70%, both methods gave comparable results with similar sequence coverage and 100% concordance for lineage assignment. Further investigation revealed a higher percentage of reads aligning to the SARS-CoV-2 genome with a greater depth of coverage using the one-step method compared to the two-step method. Finally, analysis of scaled one-step and library reaction volumes revealed significant cost savings for samples undergoing WGS. Overall, the results presented here verify the accuracy and reproducibility of one-step targeted amplification and offer an efficient and cost-effective workflow for routine SARS-CoV-2 genomic surveillance.

Humans

Microbial signal profiles and organism-level concordance between plasma metagenomic sequencing and blood culture in suspected bloodstream infection.

Plasma metagenomic next-generation sequencing (mNGS) and blood culture detect different components of the microbial signal and frequently produce discordant organism reports. We characterized microbial signal class, report-derived burden, organism-level concordance, and independent clinical attribution in a retrospective, single-center, episode-level cohort. Among 329 episodes with evaluable plasma mNGS reports, 315 had blood culture performed; 232 were mNGS positive/culture negative and 53 were positive by both methods. In the 232 discordant episodes, the recorded routine-care diagnosis classified 124 as bloodstream infection (BSI) and 108 as non-BSI. Nonviral signals were present in 78.2% and 42.6%, respectively (P&#x2009;<&#x2009;0.001), and median maximum report-derived sequence counts were 98.5 and 11.5 (P&#x2009;<&#x2009;0.001). Two laboratory physicians then independently reviewed source records using structured criteria while masked to the recorded BSI label and mNGS organism and sequence-count information. Initial agreement for the five-category BSI assessment was 97.6% (Cohen's kappa, 0.960). Within the mNGS-positive/culture-negative subgroup, adjudicated BSI likelihood showed a modest ordinal association with report burden (Spearman rho&#x2009;=&#x2009;0.190; P&#x2009;=&#x2009;0.004), while mNGS organisms were considered supported in 1 episode, plausible in 158, unlikely or contaminant in 72, and unresolved in 1. Among 53 dual-positive episodes, 33 (62.3%) shared at least one species, but only 5 (9.4%) had complete species-set concordance. Plasma mNGS and blood culture therefore frequently generated non-equivalent organism sets. Signal class and report burden contributed graded contextual evidence, but organism-level attribution required clinical review and orthogonal microbiology rather than binary positivity alone.

Humans