Search PubMedSearch

SEARCH · Search PubMed

Results for “Genomic epidemiology”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

428 records · Page 8Linked to original sources

A genome-wide coverage-based pipeline for the identification of host-derived candidate DNA biomarkers from cell-free blood.

We have created a new data-analysis pipeline for the discovery of host-specific candidate DNA biomarkers derived from sequencing data of cell-free blood. Unlike approaches that rely on specific molecular or genetic signatures, our method leverages the coverage distribution of cell-free DNA sequences mapped to a reference genome, applying statistical analyses to identify informative short genomic regions for biomarker discovery. The pipeline is applicable to diverse diseases and can be used to analyze cell-free DNA sequences from plasma or serum to identify candidate biomarkers that are characteristic of disease states in mammals. Core functionalities were developed in Java and integrated with open-source software tools for the preprocessing of raw sequencing data, complemented by Python scripts for the machine-learning analysis and statistical validation. The pipeline is designed for HPC use and users can access the pipeline through a Galaxy workflow, which offers a user-friendly web interface for input selection prior to execution and analysis progress monitoring. Performance tests, carried out using duplicate sets of COVID-19 samples and controls, showed linear scalability of execution time with an increasing dataset size, as well as a substantial reduction in execution time through parallelized computation, whereby each HPC node is used to process the data of one chromosome. Further statistical tests confirmed the quality of the pipeline's results by showing that the set of identified candidate biomarkers remained stable across varying dataset sizes.

Biomarkers

Integrated exome and mitochondrial genome sequencing reveals the genetic landscape of primary mitochondrial diseases: findings from a large Tunisian cohort.

Primary mitochondrial diseases are a heterogeneous group of neurometabolic disorders recognized as the most common metabolic genetic diseases. They manifest at any age, affecting any tissue or organ, especially those with high energy demands, and are caused by pathogenic variants in both mitochondrial and nuclear genomes. Here, we aimed to describe the genetic spectrum of a Tunisian pediatric cohort with suspected mitochondrial diseases. We recruited 47 unrelated families who underwent exome sequencing as a first-tier test followed by whole mitochondrial genome sequencing for unsolved cases. Dedicated bioinformatic pipelines and prediction tools were used to determine the potential disease-causing variants. Sanger sequencing confirmed the presence and segregation within parents. For the newly identified variants, structural modeling was conducted to study the impact of these variants on protein structure and motions. Dual genome sequencing yielded a molecular diagnosis in 33/47 families (70%) and 18/47 (38%) showed disease-causing variants in genes encoding mitochondrial proteins. Among them, four families disclosed novel variants in FASTKD2, SERAC1 and GATB, which were supported by in-depth in silico and structural analyses demonstrating their deleterious effect. The remaining families (32%, 15/47) disclosed other metabolic and neurological disorders. An exome-first strategy delivers a high diagnostic yield in Tunisia, where consanguinity remains high and simultaneously captures mitochondrial and non-mitochondrial etiologies. Mitochondrial sequencing remains indispensable in the case of an inconclusive exome. Thus, our data expand the clinical and genetic spectrum of primary mitochondrial diseases in Tunisia, an underrepresented and admixed population.

Humans

Genome-wide SNP data support species boundaries in sympatric Polylepis Ruiz & Pav. (Rosaceae) species from Bolivia and Ecuador.

Species delimitation in the South American genus Polylepis is notoriously challenging due to high morphological similarity and phenotypic plasticity, likely driven by hybridization and gene flow. Previous phylogenetic studies suggested that genetic structure aligns more strongly with geography than with taxonomy, questioning existing species concepts and hampering conservation efforts. We used double-digest RAD sequencing (ddRADseq) to generate genome-wide SNP data for 11 Polylepis species sampled across multiple localities in Bolivia and Ecuador. Population genetic analyses, phylogenetic inference, and network approaches were combined to assess whether genetic structure aligns more closely with taxonomy or geography. Morphologically defined species formed largely cohesive genetic lineages across regions, with species identity explaining substantially more genetic variation than locality. While localized admixture and reticulation were detected among closely related taxa, widespread species showed strong genetic cohesion and clear separation from congeners. Our results indicate that the sampled Polylepis species from Bolivia and Ecuador maintain distinct genetic identities despite localized signals consistent with gene flow. This genome-wide support for current taxonomy highlights Polylepis as a valuable model for studying speciation under gene flow and indicates that multiple geographic sampling will be essential in reconstructing a robust phylogeny of the genus, with important implications for conservation planning in Andean montane forests.

Bolivia

The impact of the COVID-19 pandemic on osteoporotic fractures: a systematic review and meta-analysis.

BACKGROUND: Recent reports suggest that the COVID-19 pandemic and associated lockdowns may have influenced the epidemiology of osteoporotic fractures, but results vary across regions and fracture types. The aim of this study was to provide evidence-based insights into the impact of the pandemic on osteoporotic fracture incidence. METHODS: We searched four databases (PubMed, Embase, Cochrane Library, and Web of Science) up to August 2025 for observational or retrospective studies comparing osteoporotic fracture incidence during the COVID-19 pandemic (2020) with the pre-pandemic period (2019). The primary outcome of interest was the change in fracture incidence, analysed using risk ratios (RR) with 95% confidence intervals (CI) in Review Manager 5.4. Subgroup analyses were performed by sex, geographic region, and fracture type. RESULTS: Nine studies meeting the inclusion criteria were analysed. Overall, "all types" of osteoporotic fractures showed a significant decrease during the pandemic (RR = 0.85, 95% CI 0.80-0.91, p&#x2009;<&#x2009;0.0001). Specifically, forearm fractures decreased significantly (RR = 0.87, 95% CI 0.79-0.96, p&#x2009;=&#x2009;0.002). However, for the most clinically significant fractures, no statistically significant global change was found for hip fractures (RR = 0.93, 95% CI 0.76-1.15, p&#x2009;=&#x2009;0.14) or vertebral fractures (RR = 1.35, 95% CI 0.85-2.15, p&#x2009;=&#x2009;0.20). In regional subgroup analysis, hip fracture incidence decreased significantly in South America (RR = 0.79, p&#x2009;=&#x2009;0.0004) and in both males and females, but no significant change was observed in Europe (RR = 0.92, 95% CI 0.81-1.04, p&#x2009;=&#x2009;0.17). CONCLUSION: During the COVID-19 pandemic, there was a decrease in the incidence of minor fractures, such as those of the forearm, likely due to reduced outdoor activity. However, the incidence of major osteoporotic fractures (hip and vertebral) remained stable globally, with significant reductions observed only in specific regions like South America.

Humans

Genome-wide identification and functional validation of asparagine synthetase genes (NtASNs) in Nicotiana tabacum.

Asparagine (Asn) is pivotal for plant nitrogen (N) metabolism and plays indispensable roles in plant growth, development, and stress tolerance. However, the systematic characteristics and core functions of asparagine synthetase genes (NtASNs) in tobacco remain unclear. Through a comprehensive genome-wide investigation, nine members of the NtASN gene family were identified. Subsequent CRISPR/Cas9-mediated knockout and overexpression assays of these NtASN genes revealed that NtASN1e, NtASN2a, and NtASN2b are the core genes responsible for Asn biosynthesis in tobacco. Their knockout reduced asparagine synthetase activity and Asn content, delayed seed germination by 2-3 days, and displayed elevated oxidative injury when exposed to salinity conditions. In contrast, overexpression of these genes elevated Asn accumulation. Subcellular localization analysis indicated that NtASN1e was localized to both the cytoplasm and chloroplasts, whereas NtASN2a exhibited dual localization in the cytoplasm and endoplasmic reticulum, and NtASN2b was mainly localized in the cytoplasm. This study systematically clarifies the evolutionary characteristics and core functions of the NtASN gene family and provides candidate genes for optimizing nitrogen metabolism and improving salt-stress adaptation in tobacco. These findings hold important practical significance for molecular breeding and product quality improvement in industrial crops.

Nicotiana

Tigecycline-resistant Staphylococcus in waiting pens of a pig slaughterhouse: genomic insights into a food safety alert.

BACKGROUND: The waiting pens of slaughterhouses represent a critical control point in the 'farm-to-fork' continuum, yet their role in the emergence and dissemination of antimicrobial resistance remains understudied. This study investigated tigecycline-resistant Staphylococcus (TRS) in these high-risk zones to assess their prevalence, resistance mechanisms, and transmission dynamics. METHODS: 400 samples were collected from the waiting pens of a pig slaughterhouse in Guangzhou, China. Antimicrobial susceptibility testing, whole-genome sequencing, phylogenetic analysis, and molecular cloning were employed to characterize resistance mechanisms and transmission patterns. RESULTS: 78 TRS strains were isolated and classified into three species, including S. borealis, S. ureilyticus, and S. pasteuri. These isolates exhibited multidrug-resistant phenotypes and carried new mutations in rpsJ and tet(M), which were functionally confirmed to reduce tigecycline susceptibility. Phylogenetic evidence demonstrated clonal transmission between pig farms and the slaughterhouse. The tet(M) gene was located within Staphylococcal cassette chromosome mec elements mediated by IS257, while tet(L) was carried by plasmids formed through IS256/IS257-mediated recombination. CONCLUSIONS: Waiting pens serve as crucial reservoirs for the amplification and dissemination of antimicrobial resistance. Our findings underscore the urgent need for enhanced biosecurity measures, improved waste management, and routine molecular surveillance in these high-risk zones to mitigate the spread of resistance along the food production chain.

Animals

Genome-wide identification and expression profiling of CSP and OBP genes in Stictocephala bisonia reveals candidate genes potentially associated with insecticide response.

Stictocephala bisonia is an important invasive agricultural pest. Due to the frequent application of insecticides in its habitat, this species is under intense selection pressure. Chemosensory proteins (CSPs) and odorant-binding proteins (OBPs) are known to play key roles in insecticide resistance, but their specific functions in S. bisonia remain unclear. In this study, we identified a total of 22 SbisCSPs and 16 SbisOBPs based on the S. bisonia genome. To screen for candidate genes potentially linked to insecticide resistance, we adopted a multi-criteria screening strategy that integrated phylogenetic analysis, molecular docking with three insecticides, and tissue-specific expression profiling. Phylogenetic analysis identified several SbisCSPs and SbisOBPs clustering with genes known to be involved in insecticide resistance, serving as an initial evolutionary filter. Molecular docking results indicated that &#x3bb;-Cyhalothrin exhibited the strong predicted binding affinity with most of SbisCSPs and SbisOBPs. Subsequent qPCR validation of seven prioritized candidates revealed distinct expression patterns: SbisCSP22 was highly expressed in adults and demonstrated strong binding affinity to all three insecticides tested, suggesting a potential role in mediating multi-insecticide response. Conversely, SbisCSP17 was significantly upregulated in larvae, clustered with genes known to mediate imidacloprid resistance, and exhibited strong binding affinity to imidacloprid. Given its larval-specific expression and the soil-dwelling behavior of larvae, we hypothesize that SbisCSP17 is a key candidate gene for larvae coping with soil-treated insecticides.

Animals

Bergamottin, a bioactive component of bergamot: dual inhibition of Japanese encephalitis virus internalization and genome replication.

Japanese encephalitis virus (JEV) is associated with high mortality and severe neurological sequelae, and existing prevention and control strategies remain insufficient. Therefore, the development of novel antiviral agents is of critical public health importance. This study systematically evaluated the antiviral activity and underlying mechanism of bergamottin, a natural product. Bergamottin exhibited significant dose-dependent inhibitory effects against JEV in multiple cell lines, including BHK-21, HuH-7, and Vero cells, demonstrating potent antiviral efficacy. Mechanistic investigations revealed that bergamottin primarily targeted the internalization and replication stages of the JEV life cycle, thereby effectively suppressing viral proliferation. Additionally, adaptive mutation screening indicated that the D389G mutation in envelope protein E confers drug resistance by potentially changing E protein conformation or reducing endocytic efficiency. In vivo experiment, bergamottin significantly reduced viral loads in mouse brain tissue and effectively improved the survival rate of infected mice. Our findings indicated that bergamottin exerted antiviral activity by dual targeting of key steps in the viral life cycle, making it a highly promising candidate for anti-JEV therapy. Further exploration of the antiviral properties of bergamottin is expected to facilitate its clinical development as a treatment for JEV infection.

Animals

Conserved host-exclusive oligonucleotide motifs enriched in pathogenic genes of human oncogenic viruses.

Comparative viral genomics can reveal sequence-level constraints influencing virus-host interactions. Relative minimal absent words (rMAWs) are short oligonucleotide motifs present in viral genomes but completely absent from the host, potentially reflecting selective pressures related to host adaptation and immune evasion. Using the EAGLE algorithm and the GRCh38 human reference genome, we systematically screened for prevalent rMAWs (prMAWs) across six major human oncogenic viruses: Epstein-Barr virus (EBV), hepatitis B virus (HBV), hepatitis C virus (HCV), human papillomavirus (HPV), human T-cell leukemia virus type 1 (HTLV-1), and human herpesvirus 8/Kaposi's sarcoma-associated herpesvirus (HHV-8/KSHV). highly conserved 11- and 12-bp prMAWs were identified in EBV, HBV, HTLV-1, and HHV-8/KSHV, with sequence prevalences ranging from 91.5% to 97.9%. Conversely, no short prMAWs were detected in HCV or HPV, likely reflecting differences in genome architecture, mutation rates, and long-term host adaptation to the human host. Importantly, the identified host-exclusive motifs exhibited non-random genomic distribution and were preferentially embedded within viral genes central to replication, persistence, immune modulation, and oncogenesis, including EBNA-1 (EBV), HBx (HBV), Tax-associated regions (HTLV-1), and lytic replication genes of HHV-8/KSHV. Notably, all detected prMAWs were enriched in GC nucleotides and exhibited marked CpG over-representation, suggesting sequence constraints associated with epigenetic regulation and viral persistence. Collectively, these highly conserved, host-exclusive signatures offer promising, candidates for sequence-directed approaches in the diagnosis, monitoring, and investigation of virus-associated cancers.

Humans

Metabolomics and genomics reveal high diversity and concentrations of cyanopeptides during a Microcystis bloom.

Cyanobacterial blooms are an immense global problem that release complex mixtures of poorly characterized biologically active cyanopeptides into freshwater. In this study, metabolomics and genomics were used to assess the diversity and concentrations of cyanopeptides during a dense Microcystis bloom during the late summer of 2023 in Lake Champlain, a large transboundary lake situated between Canada and the United States. Despite the relatively low genetic diversity of the bloom determined by 16S rRNA metabarcoding, 151 cyanopeptides were detected by non-targeted metabolomics. This represents the most recorded cyanopeptides from a single lake plankton bloom event to date. Fifty-two cyanopeptides were previously reported and 99 represent putative new structures. Standards from the microcystin, cyanopeptolin, microginin, and anabaenopeptin groups were used to either quantify or approximate respective cyanopeptide concentrations over the sampling period. Cyanopeptolins were the most diverse (n&#x202f;=&#x202f;68) cyanopeptides and the second most abundant, reaching 12,892&#x202f;&#x3bc;g/L. Microginins were the second most diverse (n&#x202f;=&#x202f;24) and reached the highest concentrations (18,262&#x202f;&#x3bc;g/L). Anabaenopeptins were the third most diverse (n&#x202f;=&#x202f;17) cyanopeptides, reaching 4,818&#x202f;&#x3bc;g/L. Only 8 microcystins were detected, reaching 4,935&#x202f;&#x3bc;g/L, where MC-LR was the dominant congener. Target cyanopeptide biosynthesis genes for microcystins (mcyE), cyanopeptolins (mcnC), anabaenopeptins (apnD), microviridins (mdnC), and aeruginosins (aerA) were also quantified using digital droplet PCR (ddPCR). The gene copy numbers for mcyE, mcnC, and apnD were highly correlated with their corresponding cyanopeptide concentrations. Overall, the studied Microcystis bloom produced a very diverse cyanopeptide mixture with high cyanopeptide concentrations including non-microcystin groups.

Microcystis

Integrated assessment of biocontrol potential and genome analysis of endophytic Bacillus velezensis MGL-B1 against mango stem-end rot.

Mango stem-end rot is a globally significant postharvest disease that severely threatens the mango industry, primarily caused by Botryosphaeria dothidea. However, information on biocontrol agents targeting this pathogen in mango remains limited. In this study, we isolated and identified a strain of Bacillus velezensis MGL-B1 from mango leaf tissues for the first time, which exhibited broad-spectrum antifungal activity. Both in vitro and in vivo assays demonstrated that MGL-B1 effectively inhibited the growth of B. dothidea, with an in vivo biocontrol efficacy reaching 83.72&#xa0;&#xb1;&#xa0;5.10%, comparable to that of the commonly used chemical fungicide thiabendazole. Further mechanistic analysis revealed that MGL-B1 acts by directly disrupting the integrity of the pathogen's mycelial cell membrane. In addition, its released volatile organic compounds (VOCs) also displayed significant antifungal activity, with components such as 2-nonanone, 2-nonanol, and phenylethyl alcohol being confirmed to exert antifungal effects in in vitro fumigation assays. qPCR analysis showed that MGL-B1 treatment significantly upregulated the transcriptional levels of genes involved in plant-pathogen interaction, phenylpropanoid biosynthesis, and antioxidant defense pathways in mango fruits, with upregulation folds of 16.32, 37.19, and 75.93, respectively; meanwhile, the expression of browning-related genes such as polyphenol oxidase (PPO) was markedly suppressed. Whole-genome sequencing further revealed 14 biosynthetic gene clusters for antimicrobial compounds, including five unknown gene clusters. Collectively, B. velezensis MGL-B1 represents a promising biocandidate strain with multiple antifungal mechanisms and excellent control efficacy, providing a valuable resource for green and sustainable management of mango diseases.

Mangifera

Genomic identification and functional characterization of the nuclear receptor gene family in relation to sex determination and gonad development in the Pacific oyster (Crassostrea gigas).

Nuclear receptors (NRs) are a large superfamily of transcription factors that control a wide range of physiological processes by modulating the expression of downstream target genes. Numerous studies have confirmed that NR family members play critical and conserved roles in sex determination and gonadal development across metazoans. However, in mollusks, systematic characterization of NRs and their potential functions in gonadal regulation remain largely unexplored. In this study, 46 NR gene family members in the Pacific oyster (Crassostrea gigas) were identified and assigned to eight subfamilies. All NR family members contain at least one of the two core domains (DNA-binding domain, DBD; ligand-binding domain, LBD), and conserved exon-intron structures were observed within the same subgroup, indicating their evolutionary conservation. Furthermore, expression profiling revealed high expression of CgNR2F, CgNR5A1-1, and CgNR0B1 in undifferentiated gonads, suggesting their potential involvement in sex determination. CgNR1A and CgNR2E5 were specifically expressed in female gonads and exhibited female-biased expression patterns, indicating a putative role in ovarian development. Moreover, CgNR3A and CgNR3B showed high expression levels during the undifferentiated stage and early male development stage, implying their possible participation in male gonadal development and gametogenesis. These results expand the understanding of the NR gene family in C. gigas and help elucidate the potential functions of NR genes in sex determination and gonadal development.

Animals

Genome-wide characterisation of the myosin light chain gene family in Chinese perch (Siniperca chuatsi) and its expression patterns in muscle fibre types and injury response.

The Class II myosin light chain (myl) genes in Chinese perch (Siniperca chuatsi) have not yet been systematically characterised, and relationships with muscle fibre specification, development, and injury-associated remodelling remain unclear. In this study, fast and slow muscle fibres were initially distinguished using myofibrillar ATPase histochemistry. Subsequently, genome-wide mining identified 16 Class II myl genes, comprising eight essential and eight regulatory light-chain subunits. Their conserved-domain features, chromosomal distribution, phylogenetic relationships and expression profiles were analysed. Transcriptomic profiling showed that summed myl transcript abundance was higher in fast muscle than in slow muscle, accounting for 67.2% of the pooled myl transcript pool across the two muscle types (paired t-test, raw P&#xa0;=&#xa0;0.036). mylpfa, myl1 and mylz3 were the major fast-muscle-associated genes, whereas myl10, myl2b and myl13 were preferentially expressed in slow muscle at the transcript level. These patterns support these genes as candidate fibre-type-associated expression markers. Developmental profiling identified stage-associated myl expression patterns, including a possible expression shift between mylpfb and mylpfa. In the descriptive injury-repair time course (d0-d7), FPKM profiles indicated that fast-muscle-associated genes (mylpfa, mylz3 and myl1) were lower at d1 and recovered by d3, whereas several slow-muscle-associated genes showed biphasic transcript-level increases. The slow-muscle-associated RLC gene mylpfb showed a delayed expression peak at d7. Notably, the embryonic isoform myl6l showed a modest increase from approximately 2 FPKM at d0 to 4-5 FPKM after injury, suggesting a possible injury-associated expression pattern that requires further validation. Together, these findings provide a genome-wide description of the Chinese perch myl gene family and identify candidate fibre-type-associated genes and descriptive injury-associated isoform expression patterns.

Animals

Bioprospecting microbial genomes to expand the biocatalytic toolbox of rubber oxygenases.

A set of rubber oxygenases was discovered through phylogenetic analysis and AI-based structural modeling of complexes of the putative enzymes with a substrate mimicking cis-1,4-polyisoprene. Sixteen candidate proteins were selected from thermophilic microorganisms, all sequence-related to the Latex clearing protein from Streptomyces sp. K30 (LcpK30). Sequence truncation and solubility tags were then evaluated to enhance protein expression, with the SUMO tag proving to be the most effective. Including LcpK30, nine heme-containing oxygenases were successfully expressed in E. coli NEB 10-beta cells, purified (35-157 mg L-1 yield) and characterized. Steady-state kinetics revealed significant rubber latex-degrading properties for six of them, with the truncated SUMO-fused LcpK30 (SUMO-LcpK30T) showing activity in agreement with literature. Notably, the catalytic efficiencies of all the expressed homologs lay within one order of magnitude and the oxygenase from Thermomonospora echinospora was found to be particularly promising in terms of activity, especially at high latex concentrations (more than 1% w/v). The analysis of reaction mixtures by both HPLC and HPLC-MS confirmed the oxidation of cis-1,4-polyisoprene to form the expected isoprenoid oligomers (n&#x202f;=&#x202f;2-12), whose distribution was consistent with the usual endo-type cleavage pattern in all but one case. This bioprospecting effort afforded a platform of new rubber-degrading enzymes with diverse efficiencies and product profiles, capable of adapting to targeted applications.

Oxygenases

Genome-wide identification of CXE gene family in soybean and functional characterization of GmCXE31 in lipid biosynthesis and salt tolerance.

GmCXE31 negatively regulates salt tolerance and lipid synthesis in soybean, and the cxe31-edited lines improve soybean yield and seed quality. Carboxylesterases (CXEs), as essential lipid hydrolases of the &#x3b1;/&#x3b2;-hydrolase fold superfamily, are critical for plant stress responses, hormone signaling and secondary metabolism. The key candidate gene GmCXE31 was previously identified in our laboratory through a genome&#x2011;wide association study (GWAS) of soybean lipid&#x2011;related traits. In the present study, we further identified 60 GmCXE family genes in soybean. Phylogenetic analysis clustered them into 11 conserved subfamilies. Cis-acting element analysis showed their promoters are enriched with elements related to abiotic stress, growth and hormone signaling, suggesting potential roles in soybean development and stress adaptation. GmCXE31 is highly expressed in seedling roots and responsive to strigolactones (SLs) and salt stress. Functional assays revealed that GmCXE31 negatively regulates soybean salt tolerance: its overexpression reduced salt tolerance in Arabidopsis and soybean under 150&#x202f;mM NaCl stress, while its knockout enhanced this trait. Lipid profiling revealed GmCXE31-edited lines had higher seed oil content, elevated oleic/linoleic acid ratio and lower saturated fatty acid proportion, which was achieved by regulating lipid synthesis-related genes like GmNFYA. Agronomic trait analysis showed GmCXE31-edited lines had increased nodule number, plant height and single-plant yield at maturity, with opposite phenotypes in overexpression lines. In conclusion, this study elucidates the multifaceted roles of GmCXE31 in coordinating soybean salt tolerance, lipid metabolism and agronomic traits, providing theoretical and genetic resources for salt-tolerant and high-quality soybean molecular breeding.

Glycine max

Fructophilic lactic acid bacteria as a window into multi-scale convergent evolution.

Fructophilic lactic acid bacteria (FLAB) are a group of lactic acid bacteria with unique growth characteristics, that is, poor growth on glucose. Their growth is enhanced in the presence of fructose or external electron acceptors. These organisms inhabit fructose-rich environments such as flowers, fruits, and pollinating insects, particularly honey bees. Apilactobacillus spp. and Fructobacillus spp. are representatives of FLAB, although they belong to phylogenetically distant clades. These organisms commonly possess markedly small genomes with a low number of coding DNA sequences. Furthermore, their genomes are characterized by a markedly reduced number of genes involved in carbohydrate transport and metabolism. Genome reduction in FLAB reflects convergent adaptation to fructose-rich environments rather than general genome streamlining. The two distinct FLAB genera, Fructobacillus and Apilactobacillus, independently lost more than 100 genes in statistically similar orders. In contrast, genes involved in carbohydrate and amino acid metabolism exhibited reversed orders of loss between the two genera. Furthermore, FLAB genomes lack an intact bifunctional alcohol/aldehyde dehydrogenase gene (adhE), which causes their poor growth on glucose. A comparative genomic study suggested the evolutionary process underlying adhE gene decay during adaptation to the fructose-rich environments, including pollinating insects. In conclusion, FLAB represent a unique example of habitat-driven convergent reductive evolution that can be investigated across multiple biological scales - from individual genes to whole genomes - in the diverse LAB group with a wide range of habitats, and partially share the fructophilic evolution with eukaryotic yeasts found in fructose-rich habitats.

Fructose

Emerging techniques of CRISPR/Cas system in antiviral therapy and diagnostics: Applications, limitations, and translational perspectives.

The CRISPR/Cas (clustered regularly interspaced short palindromic repeats) system is a versatile technology for developing antiviral medicines and editing viral genomes in both diagnostics and vaccine synthesis. Emerging insights into class 2 effectors, such as Cas9, Cas12, and Cas13, which target viral DNA and RNA, have revolutionized vaccines against viruses such as HIV, HPV, HBV, and EBV. Innovative diagnostic techniques such as SHERLOCK, DETECTR, and FELUDA have demonstrated system's diversity and accuracy in detecting the virus markers, supporting clinical decision-making, indicating adaptability and precision of CRISPR. This review critically evaluates CRISPR's role in RNA editing, emphasizing its importance for functional genomics and development of recombinant vaccines. Translational challenges are critically discussed, including off-target effects, delivery limitations, and ethical issues, for which unique approaches such as high-fidelity Cas variants, non-viral delivery systems, and bioethical frameworks are evaluated to address these limitations. This review also covers other social implications, such as accessibility and biosecurity risks, associated with CRISPR technologies Collectively, these advances underscore the transformative potential of CRISPR technologies in shaping next-generation antiviral diagnostics and therapeutics.

CRISPR-Cas Systems

A conserved distal-tail helical extension defines a tailspike attachment architecture in Gram-negative siphophages.

Rapid growth of bacteriophage genome collections has outpaced functional annotation of tail-tip proteins, limiting comparative analysis of host-recognition structures. Starting from a shared distal-tail gene organization in the Salmonella phages 9NA and Jersey, I developed a morphogenetic bioinformatic framework integrating gene synteny, sequence comparison, profile hidden Markov model (HMM) screening, structural evidence, structure-aware searching, and AlphaFold modeling. Comparison with the experimentally characterized lambda and Sf11 tail assemblies identified a predominantly alpha-helical C-terminal extension of the distal-tail (DT) protein associated with tailspike attachment, termed the distal-tail helical extension (DT-helix). Screening 541,986 proteins from 5167 complete NCBI RefSeq tailed-phage genomes, followed by evidence-based evaluation of sequence, genomic context, and structural architecture, identified 165 curated DT-helical-extension-associated phages. Their DT proteins segregated into six sequence groups. In the four principal multi-member groups, cognate tailspikes showed group-specific conservation in proximal N-terminal regions but substantially greater downstream diversity, consistent with sequence constraint at the DT-tailspike attachment boundary. A complementary ProstT5/Foldseek search supported the established groups but revealed no convincing additional highly divergent family. Together with the experimentally characterized Sf11 attachment interface, these findings define a recurrent morphogenetic architecture linking conserved distal-tail scaffolds to more variable receptor-binding proteins across siphophages infecting Gram-negative bacteria. Although universal exchangeability is not established, the identified scaffold-receptor-binding boundaries provide a framework for molecular characterization and rational phage engineering. Accession-level information for the 165 curated phages is available through PhageTailDB.

Viral Tail Proteins