Search PubMedSearch

SEARCH · Search PubMed

Results for “Bioinformatics analysis”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

1,492 recordsLinked to original sources

Integrated bioinformatics analysis reveals cross-talking hub genes and therapeutic agents between sepsis and acute myocardial infarction.

BACKGROUND: Sepsis and acute myocardial infarction (AMI) are two significant diseases that may share overlapping etiological mechanisms. This study aims to systematically identify core genes common to both conditions and to explore their potential as therapeutic targets and drug candidates through an integrative analysis of clinical data and bioinformatics. METHODS: The AMI dataset was obtained from the GEO database, and RNA sequencing data were collected from blood samples of patients with sepsis at our hospital. Common genes were identified using differential expression gene analysis (DEG) and weighted gene co-expression network analysis (WGCNA). Functional enrichment analyses, including Gene Ontology (GO) and Kyoto Encyclopedia of Genes and Genomes (KEGG) pathway analysis, were performed. A protein-protein interaction (PPI) network was constructed, and hub genes were identified using the MCC/Degree algorithm. Diagnostic value was assessed via receiver operating characteristic curve analysis. Immune infiltration patterns, single-cell sequencing data, and molecular docking simulations were employed to evaluate immune relevance and identify potential therapeutic compounds. RESULTS: A total of 417 genes were identified between sepsis and AMI, with enrichment analysis revealing significant involvement in inflammatory responses. Three hub genes-JAK2, MYD88, and TIMP1-were selected for further investigation. ROC curves confirmed their strong diagnostic performance for both diseases. Immune infiltration analysis showed that these core genes were significantly correlated with the infiltration levels of various immune cell types. Molecular docking indicated that quercetin exhibited stable binding affinity with the proteins encoded by these genes. qPCR validation further confirmed the upregulation of these three genes, supporting the anti-inflammatory effects of quercetin as a potential targeted therapy. CONCLUSION: JAK2, MYD88, and TIMP1 were identified as shared core genes in sepsis and AMI. These genes not only serve as potential diagnostic biomarkers but also offer novel targets for developing common therapeutic strategies for both conditions. Furthermore, quercetin emerges as a promising candidate for targeted treatment.

Humans

Community-driven advances in computational mass spectrometry: The perspective of EuBIC-MS members.

Advances in data acquisition, artificial intelligence, and integrative bioinformatics are driving the rapid evolution of computational mass spectrometry, and in turn, transforming modern proteomics, metabolomics, and lipidomics. These developments have greatly increased the scale and complexity of mass spectrometry data, underscoring the importance of evolving accurate, transparent, efficient and reproducible data processing workflows. Addressing these challenges requires collaborative innovation that brings together expertise in software engineering, statistics, and biology. The European Bioinformatics Community for Mass Spectrometry (EuBIC-MS), an initiative of the European Proteomics Association (EuPA), fosters a culture of open, community-driven development through its biennial Developers Meetings and Winter Schools. This commentary summarizes the scientific background and outcomes of the EuBIC-MS Developers Meeting 2025, which took place in Novacella, Italy. Three keynote presentations highlighted major frontiers in the field: deep proteome and phosphoproteome profiling, text mining for protein-protein interaction extraction, and scalable proteomics for AI-driven drug discovery. Seven community-selected hackathons addressed emerging challenges such as single-cell proteomics data analysis, FAIR metadata extraction, deep learning frameworks, R-Python interoperability, and DIA validation. Together, these efforts demonstrate the potential for scientific and technical innovation to arise from open collaboration, and highlight how community-driven initiatives can accelerate progress in computational mass spectrometry. SIGNIFICANCE: Modern proteomics increasingly depends on computational advances to translate complex, high-dimensional data into biological knowledge. The EuBIC-MS Developers Meeting 2025 exemplifies how community-driven collaboration can directly accelerate this process by bringing together experts from bioinformatics, statistics, and experimental proteomics to co-develop open, interoperable, and reproducible analytical tools. By fostering shared software frameworks, transparent benchmarking, and collaborative problem solving, the EuBIC-MS community helps ensure that technological innovation translates into reliable biological insights. This collaborative model strengthens the foundation for quantitative, system-level understanding of proteomes and establishes a sustainable path for integrating artificial intelligence and next-generation data acquisition into routine biological discovery. This commentary shows some current highlights in the field of computational mass spectrometry and community-based approaches undertaken during the most recent Developers Meeting to solve these challenges. The approaches discussed and initiated during the meeting - ranging from deep proteome profiling and phosphosite mapping to text mining, single-cell data analysis, and FAIR metadata extraction - address key bottlenecks that currently limit the biological interpretability and comparability of proteomics data.

Mass Spectrometry

A conserved distal-tail helical extension defines a tailspike attachment architecture in Gram-negative siphophages.

Rapid growth of bacteriophage genome collections has outpaced functional annotation of tail-tip proteins, limiting comparative analysis of host-recognition structures. Starting from a shared distal-tail gene organization in the Salmonella phages 9NA and Jersey, I developed a morphogenetic bioinformatic framework integrating gene synteny, sequence comparison, profile hidden Markov model (HMM) screening, structural evidence, structure-aware searching, and AlphaFold modeling. Comparison with the experimentally characterized lambda and Sf11 tail assemblies identified a predominantly alpha-helical C-terminal extension of the distal-tail (DT) protein associated with tailspike attachment, termed the distal-tail helical extension (DT-helix). Screening 541,986 proteins from 5167 complete NCBI RefSeq tailed-phage genomes, followed by evidence-based evaluation of sequence, genomic context, and structural architecture, identified 165 curated DT-helical-extension-associated phages. Their DT proteins segregated into six sequence groups. In the four principal multi-member groups, cognate tailspikes showed group-specific conservation in proximal N-terminal regions but substantially greater downstream diversity, consistent with sequence constraint at the DT-tailspike attachment boundary. A complementary ProstT5/Foldseek search supported the established groups but revealed no convincing additional highly divergent family. Together with the experimentally characterized Sf11 attachment interface, these findings define a recurrent morphogenetic architecture linking conserved distal-tail scaffolds to more variable receptor-binding proteins across siphophages infecting Gram-negative bacteria. Although universal exchangeability is not established, the identified scaffold-receptor-binding boundaries provide a framework for molecular characterization and rational phage engineering. Accession-level information for the 165 curated phages is available through PhageTailDB.

Viral Tail Proteins

Integrated exome and mitochondrial genome sequencing reveals the genetic landscape of primary mitochondrial diseases: findings from a large Tunisian cohort.

Primary mitochondrial diseases are a heterogeneous group of neurometabolic disorders recognized as the most common metabolic genetic diseases. They manifest at any age, affecting any tissue or organ, especially those with high energy demands, and are caused by pathogenic variants in both mitochondrial and nuclear genomes. Here, we aimed to describe the genetic spectrum of a Tunisian pediatric cohort with suspected mitochondrial diseases. We recruited 47 unrelated families who underwent exome sequencing as a first-tier test followed by whole mitochondrial genome sequencing for unsolved cases. Dedicated bioinformatic pipelines and prediction tools were used to determine the potential disease-causing variants. Sanger sequencing confirmed the presence and segregation within parents. For the newly identified variants, structural modeling was conducted to study the impact of these variants on protein structure and motions. Dual genome sequencing yielded a molecular diagnosis in 33/47 families (70%) and 18/47 (38%) showed disease-causing variants in genes encoding mitochondrial proteins. Among them, four families disclosed novel variants in FASTKD2, SERAC1 and GATB, which were supported by in-depth in silico and structural analyses demonstrating their deleterious effect. The remaining families (32%, 15/47) disclosed other metabolic and neurological disorders. An exome-first strategy delivers a high diagnostic yield in Tunisia, where consanguinity remains high and simultaneously captures mitochondrial and non-mitochondrial etiologies. Mitochondrial sequencing remains indispensable in the case of an inconclusive exome. Thus, our data expand the clinical and genetic spectrum of primary mitochondrial diseases in Tunisia, an underrepresented and admixed population.

Humans

Genome-wide identification and functional validation of asparagine synthetase genes (NtASNs) in Nicotiana tabacum.

Asparagine (Asn) is pivotal for plant nitrogen (N) metabolism and plays indispensable roles in plant growth, development, and stress tolerance. However, the systematic characteristics and core functions of asparagine synthetase genes (NtASNs) in tobacco remain unclear. Through a comprehensive genome-wide investigation, nine members of the NtASN gene family were identified. Subsequent CRISPR/Cas9-mediated knockout and overexpression assays of these NtASN genes revealed that NtASN1e, NtASN2a, and NtASN2b are the core genes responsible for Asn biosynthesis in tobacco. Their knockout reduced asparagine synthetase activity and Asn content, delayed seed germination by 2-3 days, and displayed elevated oxidative injury when exposed to salinity conditions. In contrast, overexpression of these genes elevated Asn accumulation. Subcellular localization analysis indicated that NtASN1e was localized to both the cytoplasm and chloroplasts, whereas NtASN2a exhibited dual localization in the cytoplasm and endoplasmic reticulum, and NtASN2b was mainly localized in the cytoplasm. This study systematically clarifies the evolutionary characteristics and core functions of the NtASN gene family and provides candidate genes for optimizing nitrogen metabolism and improving salt-stress adaptation in tobacco. These findings hold important practical significance for molecular breeding and product quality improvement in industrial crops.

Nicotiana

Meta-PseU: A meta-classifier for robust prediction of RNA pseudouridine modification sites from long sequences.

BACKGROUND AND OBJECTIVES: Pseudouridine (Ψ) represents one of the most abundant and conserved RNA modifications. Ψ provides an additional hydrogen-bond donor that enhances RNA structural stability and modulates translation. It participates in diverse biological processes, including RNA-protein interactions, splicing, translational control, and stress responses. Aberrant pseudouridylation is implicated in cancer, neurodegenerative disorders, and autoimmune diseases. Despite its biological importance, experimental identification of Ψ sites remains time-consuming and costly, limiting the feasibility of transcriptome-wide profiling. Computational approaches have therefore become essential complements to experimental techniques. However, state-of-the-art machine-learning and deep-learning predictors often suffer from limited generalizability due to small training datasets. To overcome these issues, we aim at constructing new long-sequence datasets and developing a novel Ψ site predictor. METHODS: New long-sequence datasets were constructed as benchmarks for RNA Ψ-site prediction. The Ψ modification sites in RMBase 3.0 were mapped to the reference genomes across three species of human, mouse, and yeast, and the RNA sequences with a length of 201 were generated by extending the upstream and downstream from the mapped, central sites. To eliminate sequence redundancy, the sequences were clustered using CD-HIT with a 70% sequence identity threshold. We developed Meta-PseU, a logistic regression-based meta-classifier that considered 118 machine learning and deep learning classifiers. The datasets and programs are freely accessible at https://github.com/kuratahiroyuki/MetaPseU. RESULTS: By optimizing model configuration, we proposed the Meta-PseU model stacking 32 machine learning and deep learning classifiers out of 118 classifiers. Meta-PseU substantially improved model generalizability, overcoming a key limitation of existing approaches. It greatly outperformed state-of-the-art predictors and achieved increasing accuracy with increasing sequence length. CONCLUSIONS: Long-sequence datasets were newly constructed as benchmarks for RNA Ψ-site prediction. Meta-PseU offers a new framework for robust Ψ-site identification by using long sequences.

Pseudouridine

[Study of a patient with azoospermia due to variant of MOV10L1 gene].

OBJECTIVE: To explore the clinical and genotypic characteristics of a patient with Sertoli cell-only syndrome (SCOS) due to variants of MOV10L1 gene. METHODS: A 27-year-old patient with Non-obstructive azoospermia (NOA) underwent routine semen analysis. Serum levels of follicle-stimulating hormone (FSH), luteinizing hormone (LH), progesterone (P), estradiol (E2), prolactin (PRL), and testosterone (T) were determined by chemiluminescence assays. Peripheral blood samples were collected for G-banded karyotyping analysis. Multiplex PCR fluorescence detection was used to screen for AZF gene microdeletions. Whole exome sequencing (WES) and Sanger sequencing were performed simultaneously. Testicular biopsy tissues were subjected to Hematoxylin-Eosin (HE) staining to assess seminiferous tubule cell composition, and MOV10L1 protein expression was detected by immunohistochemical staining. Bioinformatics tools were employed to predict the pathogenicity of variants and their impact on protein structure and function. This study was approved by the Medical Ethics Committee of the Guangdong Institute of Reproductive Sciences [Ethics No.: 2023(01)]. RESULTS: The patient's two semen analyses had failed to detect any sperm. Hormone tests indicated elevated FSH (22.32 mIU/mL) and PRL (397.6 mIU/mL), while T (3.68 nmol/L) and E2 (38.32 pmol/L) were reduced. Chromosomal karyotyping revealed 46,XY, and no AZF gene deletion was detected. WES and Sanger sequencing detected compound heterozygous variants of the MOV10L1 gene, including a c.345C>A (p.C115X) nonsense variant and a c.3323C>T (p.T1108I) missense variant, with the former being unreported previously. HE staining showed only Sertoli cells in the seminiferous tubules, confirming the diagnosis of SCOS. Immunohistochemical staining revealed absent MOV10L1 protein expression in the testicular tissue. Based on the guidelines from American College of Medical Genetics and Genomics (ACMG), the c.345C>A (p.C115X) was classified as a pathogenic variant (PVS1+PM2_Supporting+PP4), while the c.3323C>T (p.T1108I) was deemed variant of uncertain significance (PM2_Supporting+PP3_Supporting+PP4). Bioinformatics analysis demonstrated that c.345C>A (p.C115X) may cause premature termination of protein translation, while c.3323C>T (p.T1108I) may disrupt the hydrophobicity of the RNA helicase domain, reducing the active pocket volume and decreasing its affinity for MILI protein. CONCLUSION: This study has diagnosed a case of SCOS due to compound heterozygous variants of the MOV10L1 gene, which also enriched its mutational spectrum.

Humans

Alternative genetic codes in bacteria and archaea identified with a fast k-mer-based algorithm.

The genetic code is conserved across all domains of life and is often described as universal. Nevertheless, many exceptions to the "universal" code have now been documented, most of these through manual or semiautomated inspection of highly conserved genes. Modern bioinformatics tools improved our ability to find alternative genetic codes but remain computationally expensive, preventing widespread use on thousands of new species identified by sequencing environmental samples. Here, I report a >100-fold accelerated method for inferring the genetic code directly from assembled genomes and apply it to thousands of previously uncharacterized assemblies from archaea and bacteria. I describe three candidate genetic code variations, one of which, an alternative genetic code used by a family of Asgard archaea, is a unique example of sense codon reassignments for this domain. Identifying genetic code variations is important for understanding evolution of the standard code and improving accuracy of protein databases and open reading frame identification.

Genetic Code

Gastrointestinal digestion governs insect protein hydrolysis and predicted bioactive peptide release: Species-dependent implications for functional food applications.

This study investigates the digestion of insect proteins and the release of predicted bioactive peptides during human gastrointestinal digestion. Using the Infogest in vitro model, mealworm, cricket, and black soldier fly larvae (BSFL) proteins were digested and analyzed through discovery proteomics and bioinformatics to identify predicted bioactive peptides. Sequential windowed acquisition of all theoretical fragment ion mass spectra (SWATH-MS) quantified insect proteins including predicted bioactive peptide precursor proteins, the precursors of predicted bioactive peptides. Results indicated that gastrointestinal digestion strongly influences peptide release, with the gastric phase exhibiting a richer predicted bioactive peptide profile than the small intestinal phase. Many predicted bioactive peptides were rapidly hydrolysed under small intestine conditions, which may lead to reduced stability or diminished activity in vivo, potentially explaining why certain peptides show strong bioactivity in vitro but limited effects in vivo. Additionally, predicted bioactive peptide release varied by insect species, influenced by genetic factors and peptide abundance. These findings highlight the importance of species selection and consideration of proteolytic digestion patterns in optimizing insect-derived bioactive peptides for functional foods and nutraceutical applications.

Animals

Portable metagenomics for preventive surveillance and outbreak control in livestock and poultry: Pathogen detection, resistome profiling, and antimicrobial stewardship.

Conventional diagnostics for livestock and poultry outbreaks commonly rely on culture or targeted PCR panels, which may be too slow or too narrow to guide early control decisions. Portable metagenomics, particularly real-time nanopore sequencing, offers a route to broad pathogen detection, antimicrobial-resistance gene profiling, and outbreak investigation within an integrated workflow. This implementation-focused review evaluates how near-point-of-care metagenomics may support preventive veterinary medicine through earlier detection, surveillance, cohorting, biosecurity decisions, and antimicrobial stewardship. We synthesize sample-to-answer workflows for enteric and respiratory disease in food-producing animals, including sampling, nucleic-acid extraction, host depletion or target enrichment, library preparation, sequencing, bioinformatics, quality control, and interpretation. Applications in calf diarrhea, bovine respiratory disease, poultry outbreaks, mastitis, and resistome monitoring are considered alongside the central limitation that detection alone does not establish causation. Pathogen and resistance-gene signals must therefore be interpreted with clinical signs, lesions, epidemiology, controls, and confirmatory testing. We also propose a minimum reporting checklist, intended as a practical framework rather than a validated consensus standard. Portable metagenomics is not a replacement for conventional diagnostics, but appropriately validated workflows can reduce uncertainty during time-sensitive outbreaks and support more judicious antimicrobial use.

Animals

Assessing the Frequency of VEXAS-Related Canonical UBA1 Mutations in Myelodysplastic Syndrome Patients.

OBJECTIVES: Somatic mutations in the UBA1 gene cause VEXAS syndrome, which presents with inflammatory and hematological symptoms. Case studies show a strong overlap between VEXAS and myelodysplastic syndrome (MDS). Recognizing VEXAS is important for differential diagnosis in patients with both inflammation and MDS, as accurate identification guides treatment. The study focuses on determining how often canonical UBA1 mutations linked to VEXAS occur in MDS patients. METHODS: Patients diagnosed with MDS were enrolled in the study, and genomic DNA was isolated from bone marrow FFPE samples. Molecular analysis was performed using a specifically designed ARMS-PCR approach. Additionally, protein-protein interaction (PPI) studies combined with bioinformatic analyses were carried out to explore potential links between UBA1 and pyroptosis. RESULTS: Among the 149 MDS patients analyzed, none exhibited high-Variant Allele Frequency (VAF) the canonical UBA1 point mutations linked to VEXAS syndrome. PPI analysis revealed a possible association between UBA1 and the NLRP3 inflammasome component. CONCLUSIONS: Expanding the sample size and using targeted NGS or ddPCR would improve mutation detection sensitivity and could reveal UBA1 canonical and non-canonical variants and more accurately estimate the frequency of VEXAS-related mutations in the MDS population.

Humans

Divergent evolutionary strategies in spider venoms: A comparative proteomic profiling of four sympatric species from Yunnan.

Spider venoms comprise complex cocktails of bioactive molecules evolved for predation and defense, representing a valuable resource for biological research and pharmaceutical discovery. In this study, we performed a systematic analysis of venom gland extracts from four common spider species indigenous to Yunnan, China: Agelena limbata, Hippasa lycosina, Lycosa grahami, and Sinopoda pengi. Using an integrated transcriptomic and proteomic targeted profiling approach, we successfully annotated 141 distinct toxins. Comparative analysis revealed significant interspecific heterogeneity, suggesting distinct evolutionary trajectories and "weapon system economics." Both A. limbata and L. grahami exhibited a "peptide-dominant" profile anchored by neurotoxic peptides and isomerases, optimized for rapid chemical paralysis. In contrast, S. pengi displayed a distinct "protein-dominant" signature enriched with high-molecular-weight enzymes and CAP superfamily proteins, likely functioning to facilitate tissue degradation and toxin diffusion. Occupying an intermediate position, H. lycosina demonstrated a hybrid composition. These findings suggest that although these species share the same geographical range, their venom systems have undergone divergent evolutionary adaptations driven by specific ecological niches and hunting strategies. This study represents the first systematic proteomic characterization of these venom components, providing a valuable reservoir of molecular candidates while highlighting the bioinformatic nuances of analyzing whole-gland homogenates.

Animals

Characterization of ZIC5 expression in esophageal squamous cell carcinoma and its association with patient survival.

Esophageal squamous cell carcinoma (ESCC) is a prevalent malignancy known for its aggressive nature and poor prognosis. The present study aimed to investigate the expression levels and clinical importance of the Zic family member 5 (ZIC5) gene in ESCC. Gene expression data and survival information obtained from The Cancer Genome Atlas and Gene Expression Omnibus were utilized. In 176 patients with surgically resected ESCC, immunohistochemical analysis was conducted to validate the expression of ZIC5 protein in cancerous and adjacent tissues. The findings of the present study revealed a significant upregulation of ZIC5 in ESCC compared with normal tissues (P<0.05), which was further corroborated by immunohistochemistry exhibiting a notable association between ZIC5 expression and clinical parameters such as tumor size, invasion depth, lymph node metastasis and TNM staging (P<0.05). Survival analysis further indicated that high ZIC5 expression was an independent prognostic factor for poor outcomes in patients with ESCC (hazard ratio=1.519; 95% CI: 1.017-2.269; P<0.05). In addition, bioinformatic analyses predicted that hsa-microRNA-212-5p may regulate ZIC5 mRNA and gene enrichment analysis suggested that ZIC5 may facilitate ESCC progression through involvement in the cell cycle and DNA repair pathways. In conclusion, ZIC5 is highly expressed in ESCC and associated with a poor prognosis, indicating its potential as a therapeutic target and biomarker for ESCC management. Further studies are warranted to elucidate the precise mechanisms underlying the role of ZIC5 in ESCC progression.

ESCC

Systematic identification pepper CaE2F transcription factor reveals the role of CaDPb in drought stress response.

The EARLY 2 FACTOR (E2F) transcription factor (TF) family plays a pivotal role in regulating plant development and adaptations to environmental stresses. However, the physiological function of E2Fs in pepper (Capsicum annuum L.) are not well elucidated. In this work, we conduct a comprehensive genome-wide annotation of the E2F family within the Zunla-1 pepper genome and further explore the biological roles of CaDPb in response to drought stress. Through systematic bioinformatics analysis, we identify a total of nine CaE2F genes within the Zunla-1 genome, categorizing them into three distinct subgroups. Additionally, we discover multiple cis-regulatory elements in the CaE2F promoter regions associated with responses to plant hormones and drought stress. Public RNA-seq datasets reveal distinct expression profiles of CaE2F genes across various pepper tissues and their responses to environmental stimuli and plant hormones. Subsequently, the CaDPb gene is further functionally verified in drought response. Our findings indicate that TRV2:CaDPb silenced pepper plants are more sensitivity to drought. Furthermore, we show that CaDPb participates in the regulation of reactive oxygen species (ROS) production, the expression of drought-responsive genes, and the modulation of stomatal aperture. Taken together, our findings provide a comprehensive characterization of E2F genes in pepper and offer insights into the biological function of CaDPb in pepper drought stress response.

Capsicum

Influence of nicotine on protein expression around hydrophilic osseointegrated implants: A proteomic study in male rats.

OBJECTIVE: To ensure the success of dental implant treatment, various factors must be considered, including osseointegration and systemic conditions. There is evidence in the literature that smokers may exhibit alterations in tissue healing, which can compromise the success of implant rehabilitation. Therefore, this study aimed to investigate the influence of nicotine on the protein profile of bone tissue around hydrophilic implants during the osseointegration process in rats. DESIGN: Bone tissue samples from the control and nicotine groups (n&#x202f;=&#x202f;3 per group) were subjected to protein extraction, mass spectrometry, and bioinformatic analyses. Protein identification was performed using Proteome Discoverer 2.1 software and the SEQUEST algorithm, and the protein data were compared with those of a protein database of Rattus norvegicus obtained from UniProt. RESULTS: A total of 740 proteins were detected in both the control group and the nicotine-exposed group. Among them, the proteins biglycan, periostin and histone H4 were highlighted because of their higher abundance in the healthy implant group, while they were reduced in the nicotine-exposed group. CONCLUSIONS: Nicotine has the potential to alter the protein profile of bone tissue around hydrophilic implants during osseointegration, which may impair tissue remodeling and healing.

Animals

Metatranscriptomic analysis of viral sequences associated with Culex nigripalpus at an Alabama aquaculture site.

Mosquitoes associated with aquaculture habitats can harbor diverse viruses, yet the viromes of many locally abundant species remain poorly characterized. At an aquaculture-associated site in Auburn, Alabama, we surveyed mosquito populations and found Culex nigripalpus to be the dominant species collected. To characterize viruses associated with this mosquito, we performed RNA-seq on pooled female Cx. nigripalpus and compared complementary bioinformatic workflows for viral detection and genome recovery. One workflow removed host-associated reads by mapping to the closest available mosquito reference genome prior to assembly, whereas a second workflow used fully de novo assembly and viral database annotation. Additional protein-level filtering, cross-workflow comparison, and comparison of Trinity and rnaSPAdes assemblies were used to prioritize well-supported viral candidates. Across the original analyses, 16 submitted accessions corresponding to 12 collapsed virus/name groups were recovered, including Merida virus, Hubei mosquito virus 5, Zhejiang mosquito virus, Hubei virga-like virus 3, Rinkaby virus, Elemess virus, Qingnian mosquito virus, Serbia narna-like virus 2, XiangYun narna-levi-like virus 8, Ecclesville picorna-like virus, and baculovirus-like fragments. Several candidates were supported across multiple workflows, while others were recovered only under specific analytical conditions, indicating that candidate recovery was influenced by assembly and filtering choices. Selected viral contigs were independently supported by RT-PCR amplification. Overall, these results provide a first characterization of viral sequences associated with Cx. nigripalpus from an Alabama aquaculture-associated site and show that comparison across assembly and filtering strategies helped prioritize the most consistently supported viral candidates.

Animals

Genome-wide identification of the peanut HD-Zip gene family and AhHDZ15 positively regulating salt and drought stress in heterologously overexpressed Arabidopsis.

Homeodomain-leucine zipper (HD-Zip) transcription factors play important roles in plant growth, development, and abiotic stress responses. However, bioinformatic analyses and functional studies of HD-Zip family in peanut are scarce. In this study, 128 AhHDZ genes were identified and classified into four subfamilies in the phylogenetic analysis. Transcriptomic data and RT-qPCR analysis indicated the expression levels of AhHDZ4 and AhHDZ15 were significantly elevated in response to 12&#x202f;h of salt stress, while AhHDZ4/15/60/69/126 all showed a progressive increase over time in response to drought stress. AhHDZ15 protein was localized in the nucleus. Under salt and drought stress, the germination rates of AhHDZ15-overexpressing in Arabidopsis were significantly higher than wild-type (WT), and root lengths were also significantly longer than WT. In addition, the SOD, CAT, chlorophyll content, and Relative Leaf Water Content (RLWC) value of leaves in AhHDZ15-overexpressing lines were significantly higher than WT, while the MDA content was significantly lower than WT. The above results indicate that heterologous overexpression of AhHDZ15 enhanced salt and drought tolerance in Arabidopsis. Furthermore, AhHDZ15 could bind to the L1-box element of the AhVNI2 promoter, thereby activating AhVNI2 transcription and enhancing the expression of downstream salt stress-responsive genes. These findings implies a potential function of AhHDZ15 in peanut that requires further validation.

Arabidopsis

Investigation of Fatty Acid Metabolism-Associated Molecular CPOX and the Underlying Mechanism in Follicular Lymphoma.

Dysregulated lipid metabolism is a key driver of follicular lymphoma (FL). This study aimed to explore the lipid metabolism-related genes (LMRGs) and clarify the underlying roles and mechanisms in FL. Bioinformatics methods, including differential analysis, WGCNA, machine learning, and Mendelian randomization, were utilized to select the LMRGs in FL. Gene Ontology (GO) and Kyoto Encyclopedia of Genes and Genomes (KEGG) analyses were conducted to investigate the function of the key LMRG. Receiver operator characteristic (ROC) was used to evaluate the diagnostic value of the key gene CPOX. A pan-cancer analysis investigated CPOX's expression level and immune correlations. In vitro experiments using FL cell lines (WSU-FSCCL, DOHH2) validated CPOX expression, and CPOX knockdown in DOHH2 cells was used to assess its impact on viability, migration, invasion, and fatty acid metabolism. CPOX was confirmed to be a risk factor, significantly overexpressed in FL, and exhibited effective diagnostic ability in FL (AUC&#x2009;=&#x2009;0.731). Functional analysis linked CPOX to mitochondrial function, oxidative phosphorylation, and heme metabolic process. Pan-cancer indicated the dysregulated CPOX across multiple cancers and closely correlation with immune characteristics. Experimentally, CPOX was higher in the more invasive DOHH2 cells; and CPOX knockdown suppressed FL progression and reduced lipid droplet formation, triglyceride, total cholesterol, and free fatty acid levels. In conclusion, this study fills the gap in understanding the significance of lipid metabolism-related molecules in FL, and innovatively proposes that CPOX is a risk factor for FL. Knockdown of CPOX inhibits the FL progression, which is regulated by fatty acid metabolism.

Lymphoma, Follicular