Search PubMedSearch

SEARCH · Search PubMed

Results for “duplication”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

38 recordsLinked to original sources

Ramu stunt virus genome reveals previously unreported segments and nucleocapsid domain duplication in Mechlorovirus.

Ramu stunt virus (RmSV), a member of the genus Mechlorovirus within the family Phenuiviridae, was previously described as a six-segmented RNA virus infecting sugarcane. In this study, we re-examined type material and additional isolates using high-throughput sequencing and RT-PCR validation, revealing that RmSV possesses a nine-segmented genome, making it the largest reported in the Phenuiviridae. This expanded architecture includes duplicated RNA segments (RNA 2a and RNA 2b) encoding nucleocapsid-like proteins and two novel segments (RNA 7 and RNA 8). Comparative analysis showed that RNA 2a and 2b share about 84% amino acid identity, while RNA 5 encodes a third nucleocapsid homolog, indicating unprecedented domain redundancy. Structural modeling confirmed that all three nucleocapsid proteins maintain a conserved fold despite low sequence identity, with electrostatic mapping suggesting differential RNA-binding potential. Additionally, RNA 6 encodes a hypothetical protein structurally similar to the rice stripe virus disease-specific S-protein, implicating a role in symptom development. Transcript abundance analysis revealed RNA 6 as the most highly expressed segment across isolates. These findings revise the genomic composition of RmSV, highlight mechanisms of genome plasticity and adaptive evolution in plant-infecting bunyaviruses, and underscore practical implications for diagnostic assay design, resistance breeding, and biosecurity surveillance.

Genome, Viral

Genome-wide characterization of heat shock protein genes reveals thermal stress-responsive candidates in Litopenaeus vannamei.

Heat shock proteins (HSPs) are conserved molecular chaperones involved in protein folding, refolding, aggregation prevention, and degradation of damaged proteins. However, the genomic organization and thermal responsiveness of HSP genes in the Pacific white shrimp (Litopenaeus vannamei) remain incompletely understood. Here, we performed a genome-wide analysis of the HSP gene family and examined its phylogenetic relationships, structural features, duplication patterns, sequence variation, interaction networks, and transcriptional responses to acute heat stress. A total of 34 HSP genes were identified and classified into the HSP90, HSP70, HSP40/DNAJ, HSP60, and small HSP families. Phylogenetic, motif, gene structure, synteny, and subcellular localization analyses revealed evolutionary conservation and structural diversification among family members. Three duplicated gene pairs were identified, comprising two segmental duplications and one tandem duplication. All pairs exhibited Ka/Ks ratios below 1, consistent with purifying selection of varying strength. Sequence analysis identified 295 nonsynonymous single-nucleotide polymorphisms, of which 12 were consistently predicted to be deleterious by multiple algorithms. Protein-protein interaction analysis indicated enrichment of protein-folding and cellular stress-response functions. RT-qPCR analysis showed significant induction of HSPA4, HSP90AA1, TRAP1, BiP, and DNAJA1 after 6, 12, and 24 h of exposure to 34 °C, whereas DNAJC3 was significantly induced only at 12 h. All six genes reached their highest transcript abundance at 12 h. These findings may provide a genomic framework for HSP genes in L. vannamei and identify candidate genes and variants associated with thermal stress responses.

Animals

Comparative genomic and proteomic analysis reveals orthogroup structured evolution of tick protease inhibitors.

Protease inhibitors (PIs) play central roles in regulating endogenous proteolysis and host-parasite interactions in ticks. However, the evolutionary architecture underlying their diversification across tick lineages remains insufficiently resolved. Here, we performed a genome-wide comparative analysis of predicted proteomes from 14 tick species to systematically characterize PI repertoires. In total, 4931 putative PIs were identified and grouped into 20 families using the MEROPS classification system. Further, PI families such as Antistasin, WAP-type, and Pacifastin, which have not previously been systematically reported in tick genomes, were classified. Orthogroup inference demonstrated that PI expansion is structured at the level of evolutionary lineages rather than uniformly across families. By stratifying orthogroups according to duplication burden and taxonomic conservation, we identified a broadly conserved single-copy core under strong purifying selection. Motif level analysis of serpin reactive center loops further revealed conservation of inhibitory specificity within single copy orthogroups and diversification of key functional residues in duplication-associated lineages. Integration of secretion prediction and tissue-resolved proteomics from Hyalomma anatolicum and Rhipicephalus microplus demonstrated that evolutionary stratification is reflected at the protein level. Together, these findings provide an orthogroup-resolved evolutionary framework linking duplication dynamics, molecular evolution, and tissue-level protein deployment. This integrative approach offers a systematic basis for prioritizing conserved and diversified PI lineages for future functional and anti-tick intervention studies.

Animals

Upscaling Genotyping by Amplicon Sequencing With GBAS-GUI.

Genotyping by amplicon sequencing (GBAS) is a relatively low-cost approach for generating genotypic data compared with established genomic methods, making it highly scalable and particularly suitable for large-scale genetic monitoring projects. However, most existing analytical pipelines are either marker-specific, insufficiently scalable, or lacking efficient data management systems for the long-term integration of genotypic information, limiting the full potential of GBAS. Here, we address this gap by introducing GBAS-GUI (https://github.com/sonnenbe-dot/GBAS-GUI), a pipeline capable of generating GBAS-based genotypic data for a wide variety of loci at scale. GBAS-GUI integrates a graphical user interface with multiple checkpoints to improve accessibility and robustness. It implements multiprocessing architecture and a relational database that links genotypic data with associated sample metadata to enhance scalability and data management. The pipeline further enables marker screening through automated calculation of polymorphism information content (PIC) and implements a strategy to recover homologous genotypic information from paralogous loci with non-overlapping amplicon length ranges. Using multiple empirical datasets, we demonstrate substantial improvements in processing speed, database management and handling artefacts related to co-amplification of unspecific regions and duplicates of the same genomic region. We further show that incorporating the full sequence information captured by an amplicon increases marker information content beyond what is achievable with length-based genotyping alone and expands the analytical versatility of GBAS. Overall, GBAS-GUI provides a robust, scalable and versatile framework that unlocks the potential of GBAS for large-scale population genetic and phylogeographic studies.

Genotyping Techniques

Structural complexity and mechanistic diversity of MECOM rearrangements in myeloid neoplasms.

Rearrangements involving MECOM at chromosome 3q26.2 are recurrent in myeloid neoplasms, classically represented by inv(3)(q21q26.2) and t(3;3)(q21;q26.2), which reposition the GATA2-distal haematopoietic enhancer and drive aberrant EVI1 overexpression. However, the full structural and mechanistic diversity of MECOM rearrangements (MECOM-r) is yet to be explored. We retrospectively analysed 97 cases with cytogenetically defined MECOM-r and identified 12 with complex rearrangements using GTG-banded karyotyping and tri-colour interphase/metaphase fluorescence in situ hybridisation analyses. These 12 cases demonstrated remarkable structural heterogeneity. The abnormalities encompassed translocations, inversions, insertions, duplications, and deletions, which often coexisted within the same specimen as multiple rearranged subclones. Insertional events emerged as a distinct mechanism of MECOM activation. These encompassed insertions of MYNN and/or MECOM into chromosomes 1 and 6, insertion of chromosome 8 segment into MECOM, and inverted insertions between homologous chromosome 3 segments. Recurrent breakpoints at 3q21 across multiple cases, together with localised copy number imbalances frequently involving the MYNN and GOLIM4 loci at 3q26.2, underscore the architectural fragility of these two regions. Co-occurring abnormalities such as -5/del(5q), -7/del(7q), and TP53 loss were common, reflecting a permissive genomic background for chromosomal reassembly. Our findings expand the mechanistic landscape of MECOM-r beyond canonical inv(3)/t(3;3), establishing 3q21 and 3q26.2 as structural 'hotspots' and genomic instability hubs. Distinct from fusion-driven oncogenes such as KMT2A, MECOM activation results from enhancer hijacking and regional structural remodelling, leading to EVI1 overexpression and clonal evolution in myeloid malignancies.

Humans

Comparison of paralog identification methods and their impact on species tree topologies in target capture phylogenomics within the Sindora clade (Detarioideae: Leguminosae).

Target capture is a common method of generating high throughput DNA sequencing data for phylogenetic reconstruction of species relationships, for which single copy genes are usually most informative. However, a pervasive problem with target capture is that putatively single copy genes may in fact be paralogs resulting from gene duplication, which are problematic for phylogenetic inference because their evolutionary history may differ from the divergence history of species. Here, we use as a case study a target enrichment dataset of 88 species of Detarioideae (Leguminosae) with a focus on the Sindora clade to examine approaches for handling paralogs, including the built-in paralog handling functions in HybPiper and CAPTUS, plus subsequent steps using Putative Paralog Detection and the tree-based Yang & Smith orthology inference approach. We compare the paralogs flagged using these methods and verify their performance with BLAST mapping against a reference genome sequence of Sindora glabra, and then subsequently compare the species tree topologies produced across these methods. Our comparisons of paralogs flagged across the Sindora clade show that the Putative Paralog Detection pipeline was the most accurate in identifying paralogs in terms of its similarity to the BLAST mapping, followed by the built-in paralog identification function of CAPTUS. However, the results we recovered for the Detarioideae subfamily suggest that the largest differences in species tree topology resulted from the use of paralog-filtered alignments (such as with the Putative Paralog Detection pipeline and the Yang & Smith orthology inference approaches) rather than just by removing the sequences of identified paralogous genes. This was the true for HybPiper-assembled datasets but was not seen in CAPTUS-assembled datasets. In all comparisons, the topological differences caused by different paralog handling methods tended to be confined to clades where processes such as hybridisation and introgression are prevalent. Our study provides a roadmap to establish the best approach to identify, eliminate or separate paralogs in the absence of a chromosomally contiguous reference genome for a study group, and highlights the importance of careful data inspection and processing in addition to understanding the extent of paralogy and paralog characteristics (e.g. sequence divergence between copies) for their study group.

Phylogeny

Algae-to-host horizontal gene transfer in Paramecium bursaria is associated with host adaptation during endosymbiosis.

Paramecium bursaria maintains a stable endosymbiosis with green algae, yet the evolutionary consequences of this association remain unclear. Here, we screened the host genome for algal-derived horizontally transferred genes (HTGs) using a lineage-aware workflow designed to detect horizontal gene transfer (HGT) between two defined lineages. We identified 16 candidate HTGs, including four putative newly transferred genes and 12 homologous transferred genes, most of which were functionally associated with redox homeostasis and metabolism. Five HTGs showed symbiosis-dependent expression. RNAi knockdown of GH32s and SATs reduced host proliferation, total cell area, and motility, while GH32s knockdown also reduced endosymbiont load. Duplication patterns suggest that most transfers may have occurred after the P. bursaria lineage diverged from the sampled Paramecium species but before its lineage-specific whole-genome duplication (WGD). The HTGs also showed host-associated shifts in GC content and gene length, while representative HTGs retained conserved domains and functional motifs. Together, our results support algae-to-host HGT in P. bursaria and suggest that some transferred genes may contribute to metabolic integration during endosymbiosis.

Gene Transfer, Horizontal

Evolutionary architecture and lineage-specific diversification of Forkhead box transcription factors in Perna viridis.

The Forkhead box (Fox) transcription factors are evolutionarily conserved regulators of development, cell cycle, and apoptosis across metazoans. This study provides the first comprehensive genome-wide analysis of the Fox gene family in the Asian green mussel (Perna viridis). We identified 28 Fox genes distributed across 10 chromosomes. Comparative analysis reveals the absence of the FoxI, FoxQ1, FoxR and FoxS subfamily, consistent with other bivalves and indicative of lineage-specific gene loss during molluscan evolution. Notably, gene duplications in the FoxAB, FoxD, FoxH, FoxN1-4, FoxQ2 and FoxQD subfamilies may reflect functional diversification associated with environmental adaptation. Exon-intron structural variability, including intron loss in several paralogues, suggests structural diversification and potential regulatory variation. Phylogenetic reconstruction confirmed the monophyly of core Fox classes while highlighting divergent expansion patterns in lophotrochozoans. Selection analyses showed strong purifying selection across duplicated Fox paralogs, supporting functional conservation after lineage-specific expansion. Gene Ontology enrichment linked Fox genes to stress response, apoptosis, and transcriptional regulation. By integrating phylogenetic, structural, and transcriptomic analyses, this study provides a genomic framework for understanding Fox gene organisation, evolution, and tissue-associated expression patterns in Perna viridis and establishes a comparative resource for future functional studies in bivalves.

Animals

Structural and tissue-specific organisation of endocrine Fgf19 and Fgf21 signalling in rainbow trout.

Endocrine fibroblast growth factors (FGF19 subfamily) play a key role in regulating metabolic homeostasis in vertebrates. However, their functional diversification in salmonids remains poorly understood. In this study, we conducted an integrative characterisation of Fgf19 and Fgf21 signalling in rainbow trout (Oncorhynchus mykiss) by combining phylogenetic, structural and expression analyses. Phylogenetic analyses revealed the conservation of single fgf19 and fgf21 genes, despite the extensive expansion of receptors post-Ss4R (salmonid-specific fourth-round whole genome duplication). Structural modelling and molecular dynamics simulations demonstrated the stable interactions of both ligands to multiple Fgfr isoforms, with receptor-specific energetic profiles and conserved core interaction residues. Tissue expression profiling revealed clear differences from mammalian models, such as predominant hepatic fgf19 expression and the absence of hepatic fgf21 under basal conditions. In addition, there were complex and tissue-dependent distributions of fgfr and klotho transcripts. These findings support a receptor-driven diversification model of endocrine Fgf signalling in salmonids, suggesting enhanced endocrine plasticity associated with the retention of receptors following post-genomic duplication. Taken together, our findings provide new insights into the structural and regulatory organisation of endocrine Fgf signalling, as well as its potential role in metabolic regulation in rainbow trout.

Animals

Association between packed red blood cell transfusion and clinical deterioration in neonatal necrotizing enterocolitis: a systematic review and meta-analysis.

BACKGROUND: No systematic review has evaluated the existing evidence regarding the association between packed red blood cell (pRBC) transfusion and clinical worsening of necrotizing enterocolitis (NEC) in neonates. This systematic review and meta-analysis was conducted to address this knowledge gap. MATERIALS AND METHODS: We searched the Cochrane Library, EBSCO, Embase, Web of Science, Google Scholar, and PubMed for studies on pRBC transfusion and NEC published before May 10, 2025. Relevant articles were selected through title, abstract, and full-text screening. English-language case-control studies or cohort studies, or randomized controlled trials involving newborns with NEC that compared pRBC transfusion with no transfusion and reported changes in NEC clinical status were included. Review articles, systematic reviews, case reports, editorials, animal studies, duplicate publications, and studies with incomplete data were excluded. RESULTS: Five studies involving 971 neonates with NEC were included. The pooled analysis demonstrated a potential association between pRBC transfusion and clinical deterioration of NEC in neonates (odds ratio: 6.05, 95% confidence interval: 3.02-12.14). CONCLUSIONS: pRBC transfusion was associated with an exacerbation of NEC in neonates. However, these findings should be interpreted cautiously because of the small number of eligible studies included in this meta-analysis, and future large-scale, well-designed studies are needed to confirm the observed association.

Humans

Genome-wide characterization of the TGF-β superfamily identifies bmp15, gdf9, and gsdf as sex-biased candidate regulators of gonadal differentiation in the synchronous hermaphrodite Plectropomus leopardus.

The transforming growth factor-β (TGF-β) superfamily plays conserved roles in vertebrate reproduction and gonadal sex differentiation. However, its genomic repertoire and sex-biased expression patterns remain unclear in the leopard coral grouper (Plectropomus leopardus), a species with synchronous hermaphroditism. Here, we performed a genome-wide identification of the TGF-β superfamily, identifying 42 genes from the chromosome-level genome. Phylogenetic and synteny analyses indicated that segmental duplication under purifying selection contributed to family expansion. Expression profiling across multiple tissues and four gonadal developmental stages (undifferentiated, 120 dph; early differentiated, 15 months; mature testis, 3 years; mature ovary, 3 years) identified eight gonad-enriched genes, among which bmp15 and gdf9 exhibited pronounced female-biased expression, with transcripts localized exclusively to the oocyte cytoplasm, particularly in stage II-III oocytes. In contrast, gsdf showed male-biased expression and was localized in spermatogenic cells of the testis. These reciprocal expression patterns indicate that bmp15/gdf9 and gsdf are candidate factors associated with gonadal sex differentiation. Our study provides the first comprehensive characterization of the TGF-β superfamily in P. leopardus and highlights bmp15, gdf9, and gsdf as candidate sex-differentiation factors in this hermaphroditic species.

Animals

A genome-wide coverage-based pipeline for the identification of host-derived candidate DNA biomarkers from cell-free blood.

We have created a new data-analysis pipeline for the discovery of host-specific candidate DNA biomarkers derived from sequencing data of cell-free blood. Unlike approaches that rely on specific molecular or genetic signatures, our method leverages the coverage distribution of cell-free DNA sequences mapped to a reference genome, applying statistical analyses to identify informative short genomic regions for biomarker discovery. The pipeline is applicable to diverse diseases and can be used to analyze cell-free DNA sequences from plasma or serum to identify candidate biomarkers that are characteristic of disease states in mammals. Core functionalities were developed in Java and integrated with open-source software tools for the preprocessing of raw sequencing data, complemented by Python scripts for the machine-learning analysis and statistical validation. The pipeline is designed for HPC use and users can access the pipeline through a Galaxy workflow, which offers a user-friendly web interface for input selection prior to execution and analysis progress monitoring. Performance tests, carried out using duplicate sets of COVID-19 samples and controls, showed linear scalability of execution time with an increasing dataset size, as well as a substantial reduction in execution time through parallelized computation, whereby each HPC node is used to process the data of one chromosome. Further statistical tests confirmed the quality of the pipeline's results by showing that the set of identified candidate biomarkers remained stable across varying dataset sizes.

Biomarkers

Utility of hearing aid use on neurocognitive functions and auditory processing: A systematic review.

PURPOSE: Hearing loss not only affects the peripheral auditory system but also leads to changes in central auditory pathways and neurocognitive skills. The use of hearing aids has shown potential benefits in mitigating the adverse effects of hearing loss. Earlier literature showed various studies assessing the impact of hearing aid use on cognitive functions and auditory processing, a comprehensive review of the existing literature is warranted. Existing literature showed heterogeneous outcomes among studies investigated the effect of Hearing aid use in combating the adverse effects of hearing loss on neurocognitive functions and auditory processing. METHOD: Databases including PubMed, SciELO, Cochrane Library, PsycINFO, Web of Science, and Index Copernicus were searched (2005-2025) using English keywords on hearing aid effects on neurocognition and auditory processing in individuals with hearing loss. Studies on effect of hearing aid use on neurocognition and auditory processing were included, while those involving cochlear implant users or other neurological disorders unrelated to hearing loss were excluded. Titles, authors and publication years were examined, and duplicates were removed. RESULTS: Total of 61 research papers were found on specific area and were included in the present review. Findings of these 61 articles have been analysed and are reported in this article. The outcome of the present study showed usefulness of hearing aid use on neurocognitive functions and auditory processing. CONCLUSION: The present literature review emphasized the importance of early rehabilitation through hearing aid use to alleviate the adverse effects of hearing loss on neurocognitive function and auditory processing.

Humans

Long-term mortality in pediatric sepsis: a systematic review and meta-analysis.

BACKGROUND: Pediatric sepsis represents a significant factor in the mortality rates among children, with survivors remaining highly fragile during the period following discharge. While in-hospital and short-term mortality have been widely studied, the long-term mortality of pediatric sepsis is not adequately synthesized or appreciated. This study aims to estimate the long-term mortality associated with pediatric sepsis, providing a basis for optimizing post-discharge surveillance and care protocols. METHODS: This systematic review and meta-analysis followed PRISMA guidelines and was registered in PROSPERO (CRD420251137504). Exhaustive searches were conducted in PubMed, Embase, the Cochrane Library, and Web of Science for studies published from the inception of each database to June 30, 2025. Studies reporting long-term mortality in pediatric sepsis patients diagnosed using international consensus criteria were included. After literature screening, long-term mortality was pooled using a random effects meta-analysis in R statistical software. RESULTS: A total of 72,065 records were identified through database searching. After removing duplicates and screening, six studies comprising 11,318 pediatric sepsis patients were included. The pooled long-term mortality in pediatric sepsis was 11% (95% CI: 7-16%), though significant heterogeneity was observed (I2 = 98.2%, p&#x2009;<&#x2009;0.001). Sensitivity analyses yielded similar results, and evidence of publication bias was limited. CONCLUSION: Long-term mortality after pediatric sepsis was 11%, highlighting the persistent risk of mortality after hospital discharge. Further high-quality longitudinal studies are required to identify modifiable risk factors and guide evidence-based follow-up and personalized care.

Humans

Efficacy of the NMIC-150 system in identifying extended-spectrum beta-lactamases in clinical isolates.

Extended-spectrum beta-lactamases (ESBLs) are significant contributors to the growing global crisis of antimicrobial resistance. This study evaluated the performance of the NMIC-150 System for susceptibility testing of third-generation cephalosporins (3GCs) and assessed whether ceftazidime-avibactam and aztreonam-avibactam could identify ESBL-producing carbapenem-resistant Enterobacterales (CREs). A total of 278 non-duplicate clinical isolates (Klebsiella pneumoniae, E. coli, and Proteus mirabilis) were analyzed. Antimicrobial susceptibility was determined using reference broth microdilution (BMD) and the NMIC-150 System. ESBL production was defined as an &#x2265;eight-fold reduction in the minimum inhibitory concentration (MIC) of 3GCs in the presence of clavulanic acid, according to CLSI criteria. Whole-genome sequencing was performed to characterize ESBL and carbapenemase genes among 3GC-resistant isolates. A Random Forest model was used to predict ESBL-producing isolates based on MIC values. The NMIC-150 System demonstrated over 90% categorical and essential agreement with BMD for ceftazidime and ceftriaxone, along with robust predictive performance via Random Forest analysis. These findings suggest that the NMIC-150 System is a reliable platform for 3GC susceptibility testing and that an &#x2265;eight-fold MIC reduction with ceftazidime-avibactam or aztreonam-avibactam may serve as a phenotypic indicator of ESBL production in CRE isolates. In conclusion, the NMIC-150 System shows potential for routine antimicrobial resistance surveillance and may facilitate the rapid identification of ESBL-producing CREs in clinical settings.

Microbial Sensitivity Tests

Genome-wide identification, structural characterization, and evolutionary analysis of growth-related gene families in African catfish (Clarias gariepinus).

The somatotropic axis encompassing growth hormone (GH), insulin-like growth factor (IGF), myostatin (MSTN), and prolactin (PRL) signalling cascades is the master regulator of somatic growth, metabolism, and development in vertebrates. African catfish (Clarias gariepinus), a commercially pivotal aquaculture species, now possesses a chromosome-level reference genome (CGAR_prim_01v2); however, a systematic, genome-wide characterization spanning all five interconnected growth-related gene families has not previously been undertaken in this species. Here, we identified and characterized 15 growth-related genes spanning gh1, ghra, ghrb, Igf1, Igf2a, Igf2b, igf1ra, Igf1rb, Igf2r, Mstna, Mstnb, prl, prlra, prlrb, and smtlb distributed across 13 chromosomes. Complete one-to-one orthology with zebrafish confirmed strong dosage-balance conservation across >120 million years of teleost divergence. Physicochemical analysis resolved a clear biochemical dichotomy between compact, basic secreted ligands (19.88-45.81&#xa0;kDa; pI up to 10.02) and large, acidic, heavily glycosylated membrane receptors (56.82-270.80&#xa0;kDa; pI 4.85-5.97). Phylogenetic analysis confirmed 3R whole-genome duplication origins for all paralog pairs, while synteny analysis revealed a disruption of the ancestral gh1-prl chromosomal block in C. gariepinus, a finding that warrants further comparative and functional investigation. This genomic atlas provides the sequence and structural information including exon-intron boundaries, domain architecture, and chromosomal coordinates needed as a prerequisite for future marker-assisted selection and CRISPR-based myostatin-editing efforts in African catfish aquaculture, though translation into applied breeding outcomes will require subsequent functional and expression studies.

Animals

Genetic Screening of Colombian Patients With Early-Onset Parkinson Disease.

BACKGROUND AND OBJECTIVES: Early-onset Parkinson disease (EOPD), defined as symptom onset before 50 years of age, accounts for approximately 10% of patients and is suggested to have a greater genetic component than typical late-onset forms of the disease. Recessive variants in PRKN, PINK1, and DJ-1, are the most common genetic cause of EOPD, however, most studies are in patients of white ancestry. This study aims to analyze genetic variants in PRKN, PINK1, and DJ-1 in Colombian patients to help address the gap in EOPD genetic research of South American populations. METHODS: We analyzed 43 unrelated patients with EOPD using Sanger sequencing for the PRKN, PINK1, and DJ-1 genes and employed multiplex ligation-dependent probe amplification to detect copy number variants. Additionally, long-read whole-genome sequencing was conducted on 3 unresolved patients with age at onset before 30 years of age (long-read sequencing [LRS] patient A-C). RESULTS: We identified known pathogenic single-nucleotide variants and copy number variants in the PRKN gene accounting for 2 patients' disease (4.6% of patients). We observed 2 pathogenic variants in PRKN (c.155delA; p.N52Mfs*29 and c.1083+1G>A) in patient 1, who reported an age at onset of 16 years. We further detected a homozygous duplication of PRKN exons 5-6 in an additional patient, age at onset of 18 years. DISCUSSION: Our study helps characterize genetic contributors to EOPD in Colombian patients, demonstrating genetic forms (PRKN, PINK1, and DJ-1) are rare. Our results highlight a need to include diverse populations in research to improve genetic understanding of disease.

Journal Article

Identification and characterization of G protein-coupled receptors in the nocturnal halictid bee Megalopta genalis.

G protein-coupled receptors (GPCRs) are one of the largest families of membrane proteins in insects, regulating vision, neural signal transduction, and various physiological behaviors. Megalopta genalis exhibits a unique facultatively eusocial lifestyle and possesses adaptations for nocturnal activity; however, its GPCR family has not yet been systematically characterized. In this study, we performed genome-wide identification, phylogenetic analysis, and expression profiling of GPCRs in M. genalis by integrating genomic annotation and transcriptomic analysis. The results showed that a total of 99 GPCRs were identified in the genome of M. genalis, which were classified into four major families. Here, we show that M. genalis has undergone lineage-specific GPCR repertoire remodeling, marked by the expansion of novel orphan receptors and the systematic loss of multiple receptor subtypes, such as the neuropeptide receptors MIP-R and NPFR. Moreover, opsins have formed a diverse array of combinations and non-GPCR odorant receptors have undergone significant expansion via tandem duplication. Together, these features may represent part of the molecular repertoire associated with the adaptation of M. genalis to a nocturnal lifestyle. Furthermore, transcriptomic analysis revealed distinct spatiotemporal expression divergence within each of the Mth/Mthl and Fz GPCR families, suggesting functional specialization across development and adult tissues. This study provides the first systematic identification and initial functional characterization of GPCRs in M. genalis, revealing an evolutionary pattern characterized by the coexistence of contraction and expansion within the GPCR family. These findings lay a foundation for further studies aimed at elucidating the roles of these GPCRs in regulating M. genalis physiology and behavior.

Animals