Search PubMedSearch

SEARCH · Search PubMed

Results for “Single nucleotide polymorphism.”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

491 records · Page 2Linked to original sources

Responding to a protracted tuberculosis outbreak: lessons from multiple rounds of investigation in a Chinese boarding school.

PURPOSE: This study analysed a multi-semester pulmonary tuberculosis (PTB) cluster outbreak in a Chinese boarding school to provide evidence for future epidemic control. METHODS: Contacts were screened via symptoms, infection tests and chest radiography. Screening expanded progressively from close contacts to same-floor contacts, then all students and staff. Whole-genome sequencing (WGS) with single nucleotide polymorphism (SNP) and bioinformatics analysis was used for lineage classification, transmission clustering (&#x2264;12 SNPs defining a cluster) and drug resistance prediction. RESULTS: From 2020 to 2022, 20 students were diagnosed with PTB, half laboratory-confirmed. Most cases clustered in class 16 and were epidemiologically linked to the primary case (case 0), who had household PTB exposure. Case 0 and case 1 had diagnostic delays exceeding 3 and 6&#xa0;months, respectively. WGS of five isolates (case 1, 3, 4, 9 and 10) collected over three semesters showed all belonged to lineage 2 and differed by &#x2264;12 SNPs, confirming the same transmission chain. The infection rate in class 16 (46.34%) was significantly higher than other case classes (19.05%) and classes without cases (8.27%) (&#x3c7;2&#xa0;=&#xa0;61.169, p&#xa0;<&#xa0;0.001). No new cases were detected during a one-year follow-up of students involved in the outbreak after the final round of screening, nor among household contacts of all cases followed up to the present. CONCLUSIONS: Lack of entry health examinations facilitated the outbreak. Delayed diagnosis, incomplete contact screening and absence of preventive treatment led to cross-semester persistence. The infection rate disparity confirms class 16 as the outbreak epicentre. Improving community case management, extending contact follow-up and enhancing cluster outbreak measures are recommended to prevent future outbreaks.

Humans

Revealing the Shared Genetic Architecture of Metabolic Dysfunction-Associated Steatotic Liver Disease-Related Traits Through Genomic Structural Equation Modeling.

Although individual traits related to metabolic dysfunction-associated steatotic liver disease (MASLD) have been investigated through large-scale genome-wide association studies (GWASs), the shared genetic susceptibility across these traits remains unclear. We therefore conducted a multivariate GWAS of key MASLD-related traits to elucidate their common genetic architecture. We applied genomic structural equation modeling to model a latent genetic factor (MASLD-F) underlying genetically correlated MASLD-related traits, leveraging their GWAS-derived genetic correlations. We then performed functional annotations, including fine-mapping, transcriptome-wide association study, and cell- and tissue-type-specific enrichment analyses, and conducted Mendelian randomization analyses to identify modifiable risk factors. Our multivariate MASLD-F GWAS identified 50 independent variants across 48 genomic loci. Transcriptomic imputation identified several MASLD-F-associated genes, including ARNTL, NPC1, BTBD10, VDAC2, TSKU, SFMBT1, and ABHD17C. We observed significant enrichment of MASLD-F-related genetic signals predominantly in brain tissues, pancreatic islets, and the adrenal gland. Additionally, six modifiable risk factors and four modifiable protective factors for MASLD-F were identified. These findings reveal a complex shared genetic architecture underlying MASLD components, thereby expanding our understanding of disease pathogenesis and providing novel insights for precision medicine and public health interventions.

Humans

Exploring sex differences in endocannabinoid system biomarkers and their relationship with antidepressant treatment outcomes in major depressive disorder: a CAN-BIND 1 secondary analysis.

BACKGROUND: Sex differences in major depressive disorder (MDD) are well documented, but it remains unclear whether sex-related variation in peripheral endocannabinoid system (ECS)-related biomarkers is detectable in MDD. OBJECTIVES: To examine baseline sex differences in ECS-related mRNA expression, DNA methylation, and single nucleotide polymorphisms (SNPs) in MDD, and associations between baseline ECS markers and antidepressant outcomes in sex-stratified analyses. METHODS: Among 178 participants with MDD from CAN-BIND-1, all received escitalopram for 8 weeks; non-responders then received adjunctive aripiprazole from Weeks 8-16.Response was defined as &#x2265;&#x2009;50% reduction in MADRS score, and remission as MADRS&#x2009;&#x2264;&#x2009;10. ANCOVAs examined baseline sex differences and sex-stratified biomarker associations with percent MADRS reduction at Weeks 8 and 16, as well as categorical response and remission outcomes. Covariates included site, baseline MADRS, age, and ethnicity. False discovery rate correction was applied. RESULTS: Baseline sex differences in methylation were observed for CACNA1H, GABRB2, MAGL, and GABRR2, though none survived correction. No baseline sex differences in mRNA expression or SNPs were detected after correction. Lower baseline DAGLA mRNA in males was associated with greater Week 8 symptom improvement (FDR corrected). This association was not observed in females. No associations with response or remission at Weeks 8 or 16 survived correction. IMPLICATIONS: Baseline sex differences in peripheral ECS-related markers were not detected in this sample. Larger studies are needed to verify whether ECS-related biomarkers, particularly DAGLA, contribute to antidepressant outcomes in a sex-specific manner.

Humans

Harnessing Polygenic Risk Scores to Refine Venous Thromboembolism Risk Stratification.

BACKGROUND: Venous thromboembolism (VTE) is a major cause of morbidity in patients of all ages. Despite growing interest in polygenic risk scores (PRS) for VTE, their utility remains understudied. Our objective was to evaluate the independent impact of a PRS on VTE susceptibility in adults and children. METHODS: We completed a retrospective, case-control study of two separate cohorts with evaluation of three VTE PRS models, with the primary analysis focused on a 293 single nucleotide polymorphism (SNP) PRS. The adult cohort included 597 VTE cases and 31&#x2009;998 controls, and the pediatric cohort included 109 cases and 448 controls, both obtained from a de-identified databank with linked genetic data. Separate adult and pediatric multivariable logistic regressions were performed to measure the association of risk factors with VTE. RESULTS: Higher PRS in adults was significantly associated with increased odds of VTE, with each 1-standard deviation increase in PRS conferring an adjusted odds ratio of 1.25 (OR&#x2009;=&#x2009;1.25, 95% CI 1.15-1.36, p&#x2009;<&#x2009;0.001). Leading risk factors for adults were cancer (OR&#x2009;=&#x2009;2.43, 95% CI: 2.04-2.89, p&#x2009;<&#x2009;0.001) and recent surgery (OR&#x2009;=&#x2009;2.16, 95% CI: 1.83-2.54, p&#x2009;<&#x2009;0.001). The standardized PRS also exhibited increased risk for VTE in children (OR&#x2009;=&#x2009;1.38, 95% CI 1.10-1.74, p&#x2009;=&#x2009;0.003). Central venous catheterization (OR&#x2009;=&#x2009;5.65, 95% CI 3.40-9.50, p&#x2009;<&#x2009;0.001) was the foremost risk factor for pediatric VTE. CONCLUSION: VTE in adults and children is multifactorial, with clinical and genome-wide risk factors contributing. PRS may serve as a valuable adjunct to clinical risk factors for VTE risk stratification.

Humans

Comparative analysis of genomic variations among different Cdo1 paralogs for salinity-adaptation in oysters.

Under rapid climate change and anthropogenic activities, oysters, a global aquaculture species, are subjected to exacerbated culturing environments, especially for those living in in-shore estuarine species, such as Suminoe oysters Crassostrea ariakensis. This study aims to investigate the molecular mechanisms of salinity adaptation of C. ariakensis. We performed an expression genome-wide association study (eGWAS) to compare genetic regulation among 5 paralogous copies of a key salinity-related gene, cysteine dioxygenase 1 (Cdo1). A total of 40 significant eSNPs with 82 adjacent eGenes were identified in 2 copies (Cdo1_26639 and Cdo1_1666). We identified only trans-eSNPs for Cdo1_26639 and more cis-eSNPs for Cdo1_1666, and different eGenes for these 2 Cdo1 copies, which indicated that the expressional regulation of these paralogs may undergo distinct pathways. We identified 3 eGenes that exhibited identical expression patterns with Cdo1_26639 and Cdo1_1666, including 6-Pgdh, Trapp and tandem copy of Cdo1_27337. The expression correlation between Cdo1 copies and eGenes was enhanced under salinity stresses, suggesting the crucial role of eGenes in regulating Cdo1's expression in response to salinity changes. Our results provide comprehensive identification and comparison of eSNPs across different paralogous copies of one gene, along with insights into the molecular mechanisms underlying salinity tolerance, and genetic markers for breeding salinity-resistant oysters.

Animals

Mapping the immune-genetic architecture of Epstein-Barr virus-related phenotypes and multiple sclerosis through a single-cell genetic framework for target prioritization and pharmacologic hypothesis generation.

BACKGROUND: Multiple sclerosis (MS) is a severe neuroinflammatory disease causing substantial long-term disability. Strong epidemiologic evidence links Epstein-Barr virus (EBV) exposure with MS risk, but genetic evidence for immune target prioritization in EBV-related phenotypes remains limited. METHODS: We integrated single-cell cis-eQTL data from 14 immune cell types with GWASs of an EBV-related clinical phenotype and MS using a single-cell Mendelian randomization framework with colocalization analyses. Candidate eGenes were evaluated in independent cohorts. For multi-SNP instruments, we performed heterogeneity, pleiotropy, MR-Egger, weighted median, mode-based, and MR-PRESSO sensitivity analyses. We also conducted phenome-wide association analyses and queried DrugBank to annotate candidate compounds targeting prioritized genes. RESULTS: We prioritized 43 immune-cell-specific candidate eGenes with convergent genetic support, including 6 for the EBV-related phenotype and 37 for MS. SERPINB1 in NK cells was associated with increased risk of the EBV-related phenotype, whereas HLA-G was associated with decreased risk. For MS, APOM and MSH5 showed protective associations, while AHI1 showed cell-type-dependent, bidirectional associations across immune lineages. Colocalization and independent cohort evaluation supported these findings. Among FDR-significant multi-SNP associations, MR-Egger intercept tests did not indicate directional pleiotropy, although a small subset showed heterogeneity or MR-PRESSO signals. Phenome-wide analyses identified no significant adverse phenotypic associations among evaluable genes at the prespecified threshold. DrugBank annotation nominated sodium nitroprusside, fasudil, artenimol, and choline as hypothesis-generating compounds for experimental follow-up. CONCLUSIONS: This study provides a single-cell genetic framework for prioritizing immune-cell-specific candidate targets for EBV-related phenotypes and MS, and nominates genetically supported targets and pharmacologic hypotheses for experimental investigation.

Humans

Shared genetic architecture and cellular convergence between female reproductive disorders and pulmonary function: a genome-wide cross-trait analysis.

Female reproductive disorders (FRDs), including polycystic ovary syndrome, endometriosis, uterine leiomyomata, and infertility, have been epidemiologically associated with impaired pulmonary function. However, it remains unclear whether this cross-organ link reflects shared genetic etiology and, if so, which cellular mechanisms mediate it. We performed a systematic genome-wide cross-trait analysis of three FRDs and lung function traits (FEV&#x2081;, FVC, FEV&#x2081;/FVC) using GWAS summary statistics from individuals of European ancestry, integrating genetic correlation, bidirectional causal inference, pleiotropy mapping, and single-cell enrichment analyses. We identified significant negative genetic correlations between FRDs and lung volume traits, most prominently for FVC (rg range: -&#x2009;0.077 to -&#x2009;0.178). Bidirectional causal analyses indicated that FRDs have a detrimental effect on lung volume, with higher FRD genetic liability associated with reduced lung volume. Cross-trait meta-analysis identified 17 pleiotropic variants across 11 loci, with the 19q13.2 (LTBP4) and 12q13.13 (HOXC6/HOXC9) loci showing strong evidence of shared causal variants. Critically, single-cell analyses revealed that shared genetic risk converged on mesenchymal lineages across organs, specifically alveolar adventitial fibroblasts in the lung and stromal/smooth muscle cells in the endometrium. Transcriptome-wide analyses further nominated the estrogen-responsive gene RERG as a convergent gene linking these conditions with lung function. Our study revealed a shared genetic architecture between female reproductive disorders and lung function traits, providing a basis for further mechanistic investigations and potential clinical evaluation. Furthermore, our findings suggest that shared fibroproliferative and hormone-responsive pathways may offer insights into the biological mechanisms underlying these conditions.

Female

Assessment of Genetic Correlations Between Tobacco or Alcohol Use and Neurodegenerative Diseases Using East Asian Genetic Ancestry Genome-Wide Association Study Results.

Alzheimer's disease (AD) and Parkinson's disease (PD) are the most prevalent late-onset neurodegenerative diseases worldwide. Both are influenced in part by genetic factors and are currently incurable. Tobacco and alcohol, the two most common substances used among the general adult population, are potential AD/PD risk factors and are also heritable. Although important progress has been made, most existing research on the genetics of AD and PD has been carried out in individuals of European genetic ancestry. Investigations in a broad range of groups are crucial to understand disease mechanisms. Given the current availability of ancestry-specific tobacco and alcohol use as well as AD and PD genome-wide association study summary statistics, we performed global and local genetic correlation analyses using East Asian datasets. Genes within the correlated genetic regions were subsequently used to identify potentially enriched biological pathways between substance use and neurodegenerative diseases. We identified a global genetic correlation between smoking cessation and PD, which we confirmed in complementary European genetic ancestry data. Gene set enrichment analyses highlighted potentially shared genetic mechanisms between breast cancer and AD, which warrants further exploration. This work aims to promote further analyses across genetic ancestry groups.

Female

Genome-wide SNP data support species boundaries in sympatric Polylepis Ruiz & Pav. (Rosaceae) species from Bolivia and Ecuador.

Species delimitation in the South American genus Polylepis is notoriously challenging due to high morphological similarity and phenotypic plasticity, likely driven by hybridization and gene flow. Previous phylogenetic studies suggested that genetic structure aligns more strongly with geography than with taxonomy, questioning existing species concepts and hampering conservation efforts. We used double-digest RAD sequencing (ddRADseq) to generate genome-wide SNP data for 11 Polylepis species sampled across multiple localities in Bolivia and Ecuador. Population genetic analyses, phylogenetic inference, and network approaches were combined to assess whether genetic structure aligns more closely with taxonomy or geography. Morphologically defined species formed largely cohesive genetic lineages across regions, with species identity explaining substantially more genetic variation than locality. While localized admixture and reticulation were detected among closely related taxa, widespread species showed strong genetic cohesion and clear separation from congeners. Our results indicate that the sampled Polylepis species from Bolivia and Ecuador maintain distinct genetic identities despite localized signals consistent with gene flow. This genome-wide support for current taxonomy highlights Polylepis as a valuable model for studying speciation under gene flow and indicates that multiple geographic sampling will be essential in reconstructing a robust phylogeny of the genus, with important implications for conservation planning in Andean montane forests.

Bolivia

Cloning of two Hsp70 genes and association analysis between SNP haplotypes and high temperature tolerance trait in red swamp crayfish (Procambarus clarkii).

Aquaculture is suffering the challenge from high temperature climate. Two Hsp70 genes, PcHsp70-1 and PcHsp70-2, as key genes involved in the high temperature tolerance of red swamp crayfish (Procambarus clarkii) were identified and cloned in this study. Their molecular features and expression patterns were characterized, revealing the distinct tissue-specific upregulation expression under high temperature stress (33&#xa0;&#xb0;C). Two SNPs, PcHsp70-1 (SNP258) and PcHsp70-2 (SNP555) were examined to associate with high temperature tolerance in three populations (n&#xa0;=&#xa0;675). The genotypes of PcHsp70-1-SNP258 (GA) and PcHsp70-2-SNP555 (TT) were significantly associated with stronger high temperature tolerance. Notably, individuals carrying the haplotype of Hap I (GG&#xa0;+&#xa0;TT) showed a survival rate exceeding 70% under high temperature stress, whereas, the Hap VIII (AA + CT) showed it at 5.2%. RNA interference of PcHsp70-1 resulted in a significant decrease expression of the gene GSH-Px and its encoding protein (glutathione peroxidase) activity, and damage in intestinal tissue under high temperature stress. The transcriptome result revealed that PcHsp70-1 participates in regulation of the pathways related to cytoskeletal construction, immune response, apoptosis, and antioxidant defense. These findings indicate that PcHsp70 genes are crucial for the cellular stress response under high temperature stress. The developed Kompetitive Allele Specific PCR (KASP) markers provide valuable tools for the marker-assisted selection of high temperature tolerant crayfish varieties, supporting the sustainable development of aquaculture under the challenge of global warming.

Animals

Accelerated Biological Aging Increases the Risk of Head and Neck Cancer: Insights From Genetic Instruments of Epigenetic Clocks.

Epigenetic clocks are robust biomarkers of biological aging and have been associated with cancer susceptibility. However, the relationship between genetically predicted epigenetic age acceleration and head and neck cancer risk remains unclear. Using a large case-control study of 2189 head and neck squamous cell carcinoma (HNSCC) cases and 2189 age- and sex-matched controls, we investigated the associations between polygenic scores (PGSs) for multiple epigenetic clocks and HNSCC risk, and evaluated their potential causal roles using two-sample Mendelian randomization (MR). Genome-wide association study (GWAS)-identified single nucleotide polymorphisms (SNPs) associated with four epigenetic clocks (HannumAge, HorvathAge, GrimAge, and PhenoAge) were used to construct clock-specific PGSs. Logistic regression models were applied to assess associations between PGSs and HNSCC risk, while MR analyses, including inverse-variance weighted (IVW), weighted median, and MR-Egger methods, were used to infer potential causal relationships. Among the 48 epigenetic clock-associated SNPs, 12 showed nominal associations with HNSCC risk, and one variant (rs2275558 in PBX1) remained significant after Bonferroni correction (OR&#x2009;=&#x2009;0.67, 95% CI: 0.60-0.76). PGSs for all four epigenetic clocks were higher in cases than in controls. In logistic regression analyses, each standard deviation increase in HannumAge PGS was associated with a 25% higher risk of HNSCC (OR&#x2009;=&#x2009;1.25, 95% CI: 1.10-1.41), whereas HorvathAge, GrimAge, and PhenoAge PGSs showed weaker positive associations (ORs ranging from 1.06 to 1.10). Individuals in the highest PGS quartile for all four epigenetic clocks exhibiting 14%-25% higher risk than those in the lower three quartiles. MR analyses supported potential causal effects of genetically predicted HannumAge (IVW OR&#x2009;=&#x2009;1.24 per SD increase, 95% CI: 1.09-1.42) and GrimAge (IVW OR&#x2009;=&#x2009;1.23 per SD increase, 95% CI: 0.98-1.56) on HNSCC risk, with consistent estimates in weighted median analyses. Our results highlight biological aging as a potential etiologic mechanism for HNSCC and suggest that epigenetic clock-related genetic profiles may improve HNSCC risk stratification.

Humans

Introgression shapes the genomic conflict landscape of Malus, providing evidence for a reticulate backbone in a woody crop lineage.

Phylogenomic discordance is widespread across plants, but its evolutionary significance is often obscured when conflict is treated primarily as analytical noise rather than as evidence of underlying processes. In woody lineages in particular, incomplete lineage sorting, introgression, and genome duplication can interact over long timescales to produce complex genomic histories that are not adequately summarized by a strictly bifurcating tree. Here, we use Malus as a model woody genus to investigate how these processes structure conflict across a genus-scale, accession-based phylogenomic framework. Using broad taxon sampling, hundreds of nuclear loci, plastid genomes, and genome-wide SNP summaries, we reconstruct a robust nuclear backbone for sampled Malus lineages and evaluate where discordance is concentrated and which processes best explain it. Nuclear analyses resolve eight major clades, whereas conflict is non-random and localized to recurrent hotspots rather than evenly distributed across the tree. Cytonuclear discordance is similarly concentrated, especially around Clade H, represented by sampled accessions of M. tschonoskii, where localized plastid-nuclear disagreement is consistent with candidate plastid capture or organellar introgression. Multiple complementary analyses further indicate that the strongest conflict is not explained by ILS alone, but instead reflects lineage-structured introgression, while polyploid complexes represent additional localized sources of evolutionary complexity. Together, these results provide evidence for a reticulate genomic backbone in Malus and show how integrating nuclear, plastid, and genome-wide conflict analyses can help distinguish background discordance from process-specific signals in woody plant radiations. Several lineage-level reticulation hypotheses identified here should now be tested with broader population-level sampling and curated reference accessions.

Malus

Unravelling Ovarian Cancer: an analysis of the Influence of LRP1 and PAI1 Genetic Variations.

To assess the potential association between LRP1 (rs715948) and PAI1 (rs2227631, rs1799889) gene variation and ovarian cancer (OC) susceptibility. This study evaluated the genotypic and allelic distributions of LRP1 gene and PAI1 gene variants using Restriction Fragment Length Polymorphism (RFLP) analysis in 134&#xa0;&#xb0;C patients and 134 healthy controls. LRP1 (rs715948) showed a significant association with OC risk. The TC genotype was (OR&#x2009;=&#x2009;3.7823, 95% CI: 2.1732-6.5825, p&#x2009;<&#x2009;0.0001), and the CC genotype has (OR&#x2009;=&#x2009;2.1613, 95% CI: 1.0054-4.6459, p&#x2009;=&#x2009;0.0484). The C allele was significantly more frequent in cases (46%) than controls (32%) (OR&#x2009;=&#x2009;1.7684, 95% CI: 1.2443-2.5133, p&#x2009;=&#x2009;0.0015). For PAI1 (rs2227631), AG and GG genotypes showed no significant association (p&#x2009;=&#x2009;0.3519 and p&#x2009;=&#x2009;0.1165, respectively). PAI1 (rs1799889) AG genotype was (OR&#x2009;=&#x2009;5.855, 95% CI: 2.4663-13.9027, p&#x2009;<&#x2009;0.0001), while GG genotype showed no significance (p&#x2009;=&#x2009;0.1025). The dominant model of LRP1, (TC&#x2009;+&#x2009;CC) and C alleles, were significantly more frequent in OC cases, indicating a potential risk factor. In contrast, the dominant models (AG&#x2009;+&#x2009;GG) and G alleles of PAI1 (rs2227631, rs1799889) showed no significance with OC susceptibility. Genetic variation in LRP1 (rs715948) significantly associated with increased OC risk, particularly the TC and CC genotypes and C allele. The C allele of this gene is key markers linked to higher OC susceptibility. Whereas in PAI1 (rs2227631, rs1799889), dominant models (AG&#x2009;+&#x2009;GG) show no significance, association suggesting a less prominent role in OC susceptibility. These findings highlight LRP1 as a potential genetic biomarker for OC risk assessment, while the role of PAI1 variants warrants further investigation in larger sample size.

Humans

Whole-exome characterization of host genetic variation in HIV-associated genes across the high-prevalence Mizo population, Northeast India.

BACKGROUND: The Mizoram state of Northeast India has one of the highest HIV prevalence rates in Asia, yet the host genetic factors influencing HIV susceptibility in this Tibeto-Burman population remain uncharacterised. METHODS: We performed whole-exome sequencing using Illumina NovaSeq 6000, mean coverage 100X on 76 HIV-negative Mizo individuals. Variants were called using GATK HaplotypeCaller v4.3 against GRCh38p14, annotated with ANNOVAR, and filtered using hard-quality thresholds (QD&#xa0;&#x2265;&#xa0;2, SOR&#xa0;&#x2264;&#xa0;3, MQ&#xa0;&#x2265;&#xa0;40, DP&#xa0;&#x2265;&#xa0;10, GQ&#xa0;&#x2265;&#xa0;20). The allele frequencies were compared against gnomAD v2.1.1 population databases. Hardy-Weinberg equilibrium was assessed using the Wigginton exact test with Bonferroni correction. RESULTS: Post-quality filtering resulted in 12,011 sample-variants across 2,821 unique positions from 36 HIV-associated loci (33 protein-coding genes, 2 chemokine ligands, and 3 lncRNA targets). Of these, 784 observations (51 unique positions) were high-impact nonsynonymous or loss-of-function variants. ADAR rs2229857 (p.K384R, NM_015840) was the most frequently observed variant (Mizo carrier frequency&#xa0;=&#xa0;0.895; 95% CI: 0.806-0.946). CXCR1 rs16858808 (p.R335C) showed the greatest population enrichment (Mizo carrier frequency&#xa0;=&#xa0;0.197; 95% CI: 0.123-0.300; 7.65-fold carrier-frequency enrichment versus gnomAD South Asian; CADD&#xa0;=&#xa0;15.60). Sixteen of 20 tested variants deviated from Hardy-Weinberg equilibrium after Bonferroni correction (p&#xa0;<&#xa0;0.0025), predominantly showing excess homozygosity consistent with the endogamous Mizo population. The protective variant CCR5-&#x394;32 was absent in all the 76 individuals tested. CONCLUSION: This first whole-exome characterization of HIV host genes in the Mizo population identifies CXCR1 rs16858808 as the most population-enriched functional variant and reveals a pervasive endogamy signature. These findings provide a population-specific genetic framework for future HIV susceptibility studies and ART pharmacogenomics research.

Humans

Integrative genomic and transcriptomic analyses identify key regulators of skin pigmentation in Larimichthys crocea.

The yellow body coloration of large yellow croaker (Larimichthys crocea) constitutes a crucial economic trait, yet its underlying genetic regulatory mechanisms remain poorly understood. This study systematically elucidated the molecular basis of body color variation by integrating genome resequencing and skin transcriptome analyses, combined with the contextual analysis of key pigmentation-related genes and phenotypic histological validation. 200 phenotyped individuals (including yellow-selected lines, F1 progeny, and normal control groups, all derived from a well-characterized aquaculture stock) identified 39 significantly associated SNPs (-log&#x2081;&#x2080;(P)&#xa0;&#x2265;&#xa0;6), mapping to multiple candidate genes. These genes were significantly enriched in pathways related to pigment deposition (GO:0033059), melanosome organization (GO:0032438), melanogenesis, and tyrosine metabolism. Cross-developmental stage transcriptome analysis revealed 2395 differentially expressed genes (DEGs). Multi-omics integration identified eight overlapping candidate genes, including tyrp1, slc45a2, oca2, and dgat2, among which tyrp1 was prioritized for in-depth validation based on its core regulatory role in eumelanin synthesis, significant SNP association signal, and consistent downregulation in transcriptomic data. Experimental validation demonstrated that the g.895C&#xa0;>&#xa0;T mutation in exon 2 of tyrp1b was strongly significantly associated with the yellow phenotype: the frequency of mutant genotypes (TT/CT) reached 92.86%in the yellow-selected group, whereas the control group exclusively exhibited the wild-type genotype (CC). qPCR confirmed significantly downregulated tyrp1b expression in the skin of yellow individuals, consistent with the transcriptome trend. Histological and stereomicroscopic observations of skin tissues further validated the physiological basis of the yellow phenotype, revealing a significant reduction in melanophore number and abnormal melanosome morphology in yellow-phenotype individuals, accompanied by increased xanthophore density. These results suggest that tyrp1b mutation is strongly associated with the yellow phenotype. However, the presence of a wild-type CC individual in the yellow group indicates that this mutation is not strictly required for yellow coloration, suggesting that other genetic or environmental factors may also contribute to the phenotype, Additionally, downregulation of the carotenoid metabolism gene bco2 coupled with upregulation of xdh, together with the functional changes of slc45a2 and oca2, may synergistically promote xanthophore pigment deposition, contributing to the yellow phenotype. As melanin synthesis in large yellow croaker relies on the conserved tyrosinase pathway and transporter proteins, mutations in associated genes (tyrp1b, slc45a2, oca2) represent a primary underlying cause for the loss of melanin-based coloration and transition to a yellow phenotype in L. crocea. These findings provide key molecular targets and a theoretical foundation for molecular breeding of body color in this species, and also enrich the understanding of xanthism regulatory mechanisms in teleosts.

Animals

Quo vadis, BGA? A collaborative EDNAP exercise on the challenges and progress in forensic biogeographical ancestry inference.

There is a broad consensus that forensic tests for the prediction of externally visible characteristics (EVC) and analysis of biogeographic ancestry (BGA) of an individual are technically reliable. However, interpretation of the results and population-specific genotype distribution patterns remains challenging. EVC and BGA analyses provide valuable information for population genetics studies and as investigative leads for criminal cases, as well as for historical and contemporary identification tests. However, inaccurate or incorrect predictions, for example, from subjective bias in the interpretations made, have the potential to misdirect police investigations. The legal situation regarding EVC and BGA testing varies by country: ranging from countries where it is explicitly prohibited, to those without specific regulations on biogeographic ancestry prediction, and others that have already enacted laws governing its use. The reluctance to utilize these analyses is not only due to legal restrictions and data protection concerns, but also to initial limited sets of sufficiently comprehensive forensic DNA assays. Forensic BGA marker panels typically contain up to &#x223c;300 SNPs. This relatively small number of genetic markers, along with limited reference population data, complicates the interpretation of results from donors of unknown origin. This paper presents the results of a collaborative EDNAP study, which, for the first time, evaluated the approach to reporting EVC and BGA data between international laboratories. For the study, DNA from nine individuals with self-reported ancestry was collected and analysed using various forensic panels differing in the number and composition of ancestry-informative markers genotyped, comprising: the Precision ID mtDNA Whole Genome Panel, the VISAGE Basic Tool and the VISAGE Enhanced Tool for Appearance and Ancestry Prediction, and the Ion AmpliSeq&#x2122; PhenoTrivium Panel. To ensure full data protection, all SNP genotypes and uniparental marker haplotypes obtained were not shared with third parties. Instead, the genetic data were analysed using a range of commonly used population analysis software packages. These analysis outcomes were then distributed to twelve European forensic laboratories (both academic and law enforcement institutions), who were asked to prepare reports based on their interpretation of the phenotypes and ancestry they inferred from the analysis data. A questionnaire sent alongside the genetic information, aimed to evaluate which difficulties were encountered by the participants in processing the BGA analysis data they were given.

Humans

Upscaling Genotyping by Amplicon Sequencing With GBAS-GUI.

Genotyping by amplicon sequencing (GBAS) is a relatively low-cost approach for generating genotypic data compared with established genomic methods, making it highly scalable and particularly suitable for large-scale genetic monitoring projects. However, most existing analytical pipelines are either marker-specific, insufficiently scalable, or lacking efficient data management systems for the long-term integration of genotypic information, limiting the full potential of GBAS. Here, we address this gap by introducing GBAS-GUI (https://github.com/sonnenbe-dot/GBAS-GUI), a pipeline capable of generating GBAS-based genotypic data for a wide variety of loci at scale. GBAS-GUI integrates a graphical user interface with multiple checkpoints to improve accessibility and robustness. It implements multiprocessing architecture and a relational database that links genotypic data with associated sample metadata to enhance scalability and data management. The pipeline further enables marker screening through automated calculation of polymorphism information content (PIC) and implements a strategy to recover homologous genotypic information from paralogous loci with non-overlapping amplicon length ranges. Using multiple empirical datasets, we demonstrate substantial improvements in processing speed, database management and handling artefacts related to co-amplification of unspecific regions and duplicates of the same genomic region. We further show that incorporating the full sequence information captured by an amplicon increases marker information content beyond what is achievable with length-based genotyping alone and expands the analytical versatility of GBAS. Overall, GBAS-GUI provides a robust, scalable and versatile framework that unlocks the potential of GBAS for large-scale population genetic and phylogeographic studies.

Genotyping Techniques

Genome-wide characterization of heat shock protein genes reveals thermal stress-responsive candidates in Litopenaeus vannamei.

Heat shock proteins (HSPs) are conserved molecular chaperones involved in protein folding, refolding, aggregation prevention, and degradation of damaged proteins. However, the genomic organization and thermal responsiveness of HSP genes in the Pacific white shrimp (Litopenaeus vannamei) remain incompletely understood. Here, we performed a genome-wide analysis of the HSP gene family and examined its phylogenetic relationships, structural features, duplication patterns, sequence variation, interaction networks, and transcriptional responses to acute heat stress. A total of 34 HSP genes were identified and classified into the HSP90, HSP70, HSP40/DNAJ, HSP60, and small HSP families. Phylogenetic, motif, gene structure, synteny, and subcellular localization analyses revealed evolutionary conservation and structural diversification among family members. Three duplicated gene pairs were identified, comprising two segmental duplications and one tandem duplication. All pairs exhibited Ka/Ks ratios below 1, consistent with purifying selection of varying strength. Sequence analysis identified 295 nonsynonymous single-nucleotide polymorphisms, of which 12 were consistently predicted to be deleterious by multiple algorithms. Protein-protein interaction analysis indicated enrichment of protein-folding and cellular stress-response functions. RT-qPCR analysis showed significant induction of HSPA4, HSP90AA1, TRAP1, BiP, and DNAJA1 after 6, 12, and 24&#xa0;h of exposure to 34&#xa0;&#xb0;C, whereas DNAJC3 was significantly induced only at 12&#xa0;h. All six genes reached their highest transcript abundance at 12&#xa0;h. These findings may provide a genomic framework for HSP genes in L. vannamei and identify candidate genes and variants associated with thermal stress responses.

Animals