Search PubMed⌕ Search

PubMed · 15223141

Coding and non-coding DNA thermal stability differences in eukaryotes studied by melting simulation, base shuffling and DNA nearest neighbor frequency analysis.

Abstract

The melting of the coding and non-coding classes of natural DNA sequences was investigated using a program, MELTSIM, which simulates DNA melting based upon an empirically parameterized nearest neighbor thermodynamic model. We calculated T(m) results of 8144 natural sequences from 28 eukaryotic organisms of varying F(GC) (mole fraction of G and C) and of 3775 coding and 3297 non-coding sequences derived from those natural sequences. These data demonstrated that the T(m) vs. F(GC) relationships in coding and non-coding DNAs are both linear but have a statistically significant difference (6.6%) in their slopes. These relationships are significantly different from the T(m) vs. F(GC) relationship embodied in the classical Marmur-Schildkraut-Doty (MSD) equation for the intact long natural sequences. By analyzing the simulation results from various base shufflings of the original DNAs and the average nearest neighbor frequencies of those natural sequences across the F(GC) range, we showed that these differences in the T(m) vs. F(GC) relationships are largely a direct result of systematic F(GC)-dependent biases in nearest neighbor frequencies for those two different DNA classes. Those differences in the T(m) vs. F(GC) relationships and biases in nearest neighbor frequencies also appear between the sequences from multicellular and unicellular organisms in the same coding or non-coding classes, albeit of smaller but significant magnitudes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Dang D Long, Ivo Grosse, Kenneth A Marx. 2004-07-01. Coding and non-coding DNA thermal stability differences in eukaryotes studied by melting simulation, base shuffling and DNA nearest neighbor frequency analysis.. https://doi.org/10.1016/j.bpc.2004.01.001

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Sp1 is essential for p16 expression in human diploid fibroblasts during senescence.

BACKGROUND: p16(INK4a) tumor suppressor protein has been widely proposed to mediate entrance of the cells into the senescent stage. Promoter of p16(INK4a) gene contains at least five putative GC boxes, named GC-I to V, respectively. Our previous data showed that a potential Sp1 binding site, within the promoter region from -466 to -451, acts as a positive transcription regulatory element. These results led us to examine how Sp1 and/or Sp3 act on these GC boxes during aging in cultured human diploid fibroblasts. METHODOLOGY/PRINCIPAL FINDINGS: Mutagenesis studies revealed that GC-I, II and IV, especially GC-II, are essential for p16(INK4a) gene expression in senescent cells. Electrophoretic mobility shift assays (EMSA) and ChIP assays demonstrated that both Sp1 and Sp3 bind to these elements and the binding activity is enhanced in senescent cells. Ectopic overexpression of Sp1, but not Sp3, induced the transcription of p16(INK4a). Both Sp1 RNAi and Mithramycin, a DNA intercalating agent that interferes with Sp1 and Sp3 binding activities, reduced p16(INK4a) gene expression. In addition, the enhanced binding of Sp1 to p16(INK4a) promoter during cellular senescence appeared to be the result of increased Sp1 binding affinity, not an alteration in Sp1 protein level. CONCLUSIONS/SIGNIFICANCE: All these results suggest that GC- II is the key site for Sp1 binding and increase of Sp1 binding activity rather than protein levels contributes to the induction of p16(INK4a) expression during cell aging.

Base Composition↗

Application of CE for determination of DNA base composition.

DNA base composition expressed as mol% of guanine plus cytosine (% GC) or GC content is a key parameter of bacterial taxonomy and genomic analyses. Direct chemical determination methods such as HPLC as well as indirect methods based on physical properties of deoxyribonucleic acid (DNA), melting point (T(m)), and buoyant density (B(d)) have been conventionally applied to determine the GC content. However, these methods require relatively large amounts of sample DNA, time, and labor. We have developed a protocol to determine the GC content by fine separation of nucleosides with CZE. Genomic DNAs with known GC content from 23 bacterial strains were determined by CE at the optimized conditions of 27 degrees C, 20 kV in 50 mM of NaHCO(3) (pH 9.0) and 70 mM SDS added. Nucleosides from <1 microg of DNA hydrolyzed with nuclease-P1 and bacterial alkaline phosphatase were separated in a 75 microm wide and 80 cm long silica capillary. The nucleoside peak areas were determined at 254 nm in less than 12 min. The CE-based determination of GC content requires only small amounts of DNA, and thus should be applicable to environmental genomics (metagenomics), as >90% of environmental micro-organisms are nonculturable and produce only small amounts of genomic DNA.

Base Composition↗