Search PubMed⌕ Search

PubMed · 9017212

Quantification of DNA patchiness using long-range correlation measures.

Abstract

We introduce and develop new techniques to quantify DNA patchiness, and to quantify characteristics of its mosaic structure. These techniques, which involve calculating two functions, alpha(l) and beta(l), measure correlations at length scale l and detect distinct characteristic patch sizes embedded in scale-invariant patch size distributions. Using these new methods, we address a number of issues relating to the mosaic structure of genomic DNA. We find several distinct characteristic patch sizes in certain genomic sequences, and compare, contrast, and quantify the correlation properties of different sequences, including a number of yeast, human, and prokaryotic sequences. We exclude the possibility that the correlation properties and the known mosaic structure of DNA can be explained either by simple Markov processes or by tandem repeats of dinucleotides. We find that the distinct patch sizes in all 16 yeast chromosomes are similar. Furthermore, we test the hypothesis that, for yeast, patchiness is caused by the alternation of coding and noncoding regions, and the hypothesis that in human sequences patchiness is related to repetitive sequences. We find that, by themselves, neither the alternation of coding and noncoding regions, nor repetitive sequences, can fully explain the long-range correlation properties of DNA.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

G M Viswanathan, S V Buldyrev, S Havlin, H E Stanley. 1997. Quantification of DNA patchiness using long-range correlation measures.. https://doi.org/10.1016/s0006-3495(97)78721-6

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Chromosome-level genome assembly of a cosmopolitan marine harmful algal bloom diatom species Chaetoceros socialis (Chaetocerotaceae).

Chaetoceros socialis is a cosmopolitan diatom species that is crucial for maintaining marine ecosystem structure and driving elemental cycles. C. socialis can form harmful algal blooms (HABs) that may cause a negative impact on the marine ecosystems. Whole-genome information for C. socialis is still unavailable, which may hinder more targeted studies on its ecological adaptive responses and evolutionary drivers. To address this gap, we employed cutting-edge genomic technologies including PacBio single-molecule real-time (SMRT) sequencing and high-throughput chromatin conformation capture (Hi-C) to achieve the first chromosome-level genome assembly of C. socialis. The assembled genome is 60.22 Mb in size with a scaffold N50 of 7.81 Mb and has been anchored to eight pseudochromosomes. A total of 13,378 protein-coding genes were predicted, of which 12,069 (90.22%) were functionally annotated. This high-quality genomic resource provides a fundamental data platform for systematically elucidating the ecological adaptation mechanisms of C. socialis.

Chromosomes↗

Chromosome-level genome assembly of the Vermilion Snapper (Rhomboplites aurorubens).

Vermilion Snapper (Rhomboplites aurorubens, Lutjanidae) inhabits deep waters (20-300 m) from North America to Brazil and supports significant commercial and recreational fisheries. Despite its economic importance, the understanding of its basic biology remains limited. Classified as Vulnerable on the Red List due to overfishing, populations have declined by over 30% in recent generations. We assembled and annotated the first chromosome-scale genome of this species by combining PacBio long reads, Illumina short reads, and Hi-C data. The resulting assembly is 987.5 Mbp, with a scaffold N50 size of 41.3 Mbp, and includes 135 contigs clustered and ordered onto 24 chromosomes with 34,496 predicted genes. The high-quality assembly and annotation contained about 98% complete and single-copy BUSCO genes. It is the most complete, chromosome-level genome assembly of an Atlantic snapper to date. The genome assembly and supporting data are valuable tools for ecological and comparative genomics studies of snappers and other valuable commercial species within the family.

Chromosomes↗

A role for TFIIIC transcription factor complex in genome organization.

Eukaryotic genome complexity necessitates boundary and insulator elements to partition genomic content into distinct domains. We show that inverted repeat (IR) boundary elements flanking the fission yeast mating-type heterochromatin domain contain B-box sequences, which prevent heterochromatin from spreading into neighboring euchromatic regions by recruiting transcription factor TFIIIC complex without RNA polymerase III (Pol III). Genome-wide analysis reveals TFIIIC with Pol III at all tRNA genes, many of which cluster at pericentromeric heterochromatin domain boundaries. However, a single tRNA(phe) gene with modest TFIIIC enrichment is insufficient to serve as boundary and requires RNAi-associated element to restrain heterochromatin spreading. Remarkably, we found TFIIIC localization without Pol III at many sites located between divergent promoters. These sites appear to act as chromosome-organizing clamps by tethering distant loci to the nuclear periphery, at which TFIIIC is concentrated into several distinct bodies. Our analyses uncover a general genome organization mechanism involving conserved TFIIIC complex.

Chromosomes↗