Search PubMedSearch

SEARCH · Search PubMed

Results for “large-scale amplifications and deletions”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

5 recordsLinked to original sources

The complete sequence of the silkworm W chromosome uncovers its rapid evolution by large-scale duplications/deletions and translocation of W-linked genes.

The complete sequence of the W chromosome, which carries feminization activity in the silkworm, is crucial for understanding the sex-determination system in Lepidoptera. However, extensive accumulation of transposons due to lack of recombination, the very rare protein-coding genes and almost no information about molecular markers has hindered full W sequencing. We report the first complete silkworm W sequence (T2T_W, 11683305 bp) obtained by combining sequencing-assembly technologies and newly developed error detection methods, evaluated with genetically mapped W-RAPD markers, W-mutants, and W-derived BAC clones. The T2T_W sequence showed that the W is composed of a massive 92% accumulation of transposons and repeat sequences, among which the main constituents are intact LTR/LINE retrotransposons indicating recent expansions. In addition to Fem clusters producing Fem piRNA (Feminizer-derived PIWI-interacting RNA), we found 26 protein-coding genes in the W sequence. These include four gene pairs encoding zinc-finger motifs designated z1:z20 and a gene encoding serine/arginine repetitive matrix protein 1-like (SRRM1-like). To identify candidate genes for female sex-determination and differentiation we also sequenced the shortest W (3.8 Mb) from a translocation mutant with feminizing activity, which harbored four conventional genes: a Fem cluster, a pair of z1:z20 isoforms, z20-S, and a SRRM1-like gene. Phylogenetic analysis revealed that z1:z20 originated from a copy of an autosomal zinc-finger gene pair, z2:z21, translocated onto the W around 2.43 Mya and subsequently amplified to yield 4 W-linked zinc-finger gene pairs. The complete W sequence revealed that large-scale deletions and amplifications played a significant role in W chromosome evolution.

Animals

Haplotype-resolved reconstruction and functional interrogation of cancer karyotypes.

Complex karyotype changes are widespread in cancer genomes. A major gap in cancer genome characterization is the resolution of rearranged chromosomes with chromosome-length continuity. Here, we describe a two-tiered approach to determine the segmental composition of rearranged chromosomes with haplotype resolution. First, we present refLinker, a bioinformatic method for robust determination of chromosomal haplotypes using cancer Hi-C data. By contrast with existing methods, refLinker is insensitive to the presence of large-scale DNA deletions, duplications, and high-level amplification in cancer genomes. Second, we demonstrate a computational strategy to determine the segmental structure of rearranged chromosomes using haplotype-specific Hi-C contacts. We apply these methods to breast cancer genomes and provide direct evidence for long-range transcriptional changes associated with rearrangements of the inactive X chromosome. Together, these results highlight refLinker's broad utility for studying the functional consequences of chromosomal rearrangements.

Humans

CNV-Finder: Streamlining Copy Number Variation Discovery.

Copy Number Variations (CNVs) play pivotal roles in the etiology of complex diseases and are variable across diverse populations. Understanding the association between CNVs and disease susceptibility is significant in disease genetics research and often requires analysis of large sample sizes. One of the most cost-effective and scalable methods for detecting CNVs is based on normalized signal intensity values, such as Log R Ratio (LRR) and B Allele Frequency (BAF), from Illumina genotyping arrays. In this study, we present CNV-Finder, a novel pipeline integrating deep learning techniques on array data, specifically a Long Short-Term Memory (LSTM) network, to expedite the large-scale identification of CNVs within predefined genomic regions. This facilitates efficient prioritization of samples for time-consuming or costly subsequent analyses such as Multiplex Ligation-dependent Probe Amplification (MLPA), short-read, and long-read whole genome sequencing. We incorporate four genes to establish our methods-Parkin (PRKN), Leucine Rich Repeat And Ig Domain Containing 2 (LINGO2), Microtubule Associated Protein Tau (MAPT), and alpha-Synuclein (SNCA)-which may be relevant to neurological diseases such as Alzheimer's disease (AD), Parkinson's disease (PD), Progressive Supranuclear Palsy (PSP), or related disorders such as essential tremor (ET). By training our models on expert-annotated samples and validating them across diverse cohorts, including those from the Global Parkinson's Genetics Program (GP2) and additional dementia-specific databases, we demonstrate the efficacy of CNV-Finder in accurately detecting deletions and duplications. Our pipeline outputs app-compatible files for visualization within CNV-Finder's interactive web application. This interface enables researchers to review predictions and filter displayed samples by model prediction values, LRR range, and variant count in order to explore or confirm results. Our pipeline integrates this human feedback to enhance model performance and reduce false positive rates. Through a series of comprehensive analyses and validations using visual inspection, MLPA, short-read, and long-read sequencing data, we demonstrate the robustness and adaptability of CNV-Finder in identifying CNVs with regions of varied size, probe density, and noise. Our findings highlight the significance of contextual understanding and human expertise in enhancing the precision of CNV identification, particularly in complex genomic regions like 17q21.31. The CNV-Finder pipeline is a scalable, publicly available resource for the scientific community, available on GitHub (https://github.com/GP2code/CNV-Finder; DOI 10.5281/zenodo.14182563). CNV-Finder not only expedites accurate candidate identification but also significantly reduces the manual workload for researchers, enabling future targeted validation and downstream analyses in regions or phenotypes of interest.

Copy Number Variation (CNV)

High-throughput method for detecting genomic-deletion polymorphisms.

DNA microarrays have been successfully used with different microorganisms, including Mycobacterium tuberculosis, to detect genomic deletions relative to a reference strain. However, the cost and complexity of the microarray system are obstacles to its widespread use in large-scale studies. In order to evaluate the extent and role of large sequence polymorphisms (LSPs) or insertion-deletion events in bacterial populations, we developed a technique, termed deligotyping, which hybridizes multiplex-PCR products to membrane-bound, highly specific oligonucleotide probes. The approach has the benefits of being low cost and capable of simultaneously interrogating more than 40 bacterial strains for the presence of 43 genomic regions. The deletions represented on the membrane were selected from previous comparative genomic studies and ongoing microarray experiments. Highly specific probes for these deletions were designed and attached to a membrane for hybridization with strain-derived targets. The targets were generated by multiplex PCR, allowing simultaneous amplifications of 43 different genomic loci in a single reaction. To validate our approach, 100 strains that had been analyzed with a high-density microarray were analyzed. The membrane accurately detected the deletions identified by the microarray approach, with a sensitivity of 99.9% and a specificity of 98.0%. The deligotyping technique allows the rapid and reliable screening of large numbers of M. tuberculosis isolates for LSPs. This technique can be used to provide insights into the epidemiology, genomic evolution, and population structure of M. tuberculosis and can be adapted for the study of other organisms.

DNA Probes

Nationwide carrier screening for congenital adrenal hyperplasia: integrated approach of CYP21A2 pathogenic variant genotyping and comprehensive large gene deletion analysis.

BACKGROUND: Congenital Adrenal Hyperplasia (CAH) due to 21-hydroxylase deficiency (21-OHD CAH) is an autosomal recessive disorder resulting from pathogenic variants in the CYP21A2 gene. The disorder exhibits variable clinical severity, with the classical form manifesting as salt-wasting crisis in neonates, while inducing ambiguous genitalia in females and precocious puberty in males through simple virilization. Identifying at-risk couples during the preconception stage holds significance for optimizing reproductive choices. METHODS: This study included 204 unrelated preconception individuals undergoing carrier screening. A robust molecular approach was devised for rapid detection of nine prevalent CYP21A2 pathogenic variants, utilizing Amplification-Refractory Mutation System (ARMS) PCR and mass spectrometry (MS) genotyping. Complementary quantitative real-time PCR (qPCR) and PCR-based Restriction Fragment Length Polymorphism (PCR-based RFLP) assays were employed for comprehensive gene deletion analysis. The concordance of pathogenic variant detection between ARMS-PCR and MS, as well as the consistency observed in molecular insights from qPCR and PCR-based RFLP, fortified the accuracy of our methodologies. RESULTS: Our combined method could detect common pathogenic variants and large gene deletions with high concordance between ARMS-PCR, MS genotyping, qPCR, and PCR-based RFLP assays. Remarkably, two carriers exhibited significant large-scale deletions, while another manifested a carrier state due to minor-scale gene conversion. The estimated carrier frequency in our cohort using these methods was approximately 1 in 65 individuals. CONCLUSIONS: The methods used for 21-OHD CAH carrier screening offer a reliable, swift, and cost-effective approach for detecting common pathogenic variants and large deletions. Despite some limitations, such as the inability to detect all rare mutations, the techniques provide a practical solution for carrier screening, with an estimated carrier frequency of 1 in 65 in our study population. These findings support the potential adoption of these methods in national carrier screening programs, offering a practical balance between efficiency and affordability.

Humans