Search PubMedSearch

Biomedical subjects

C Saccone

Publications and source records attributed to C Saccone.

At least 19 recordsLinked to original sources

WORDUP: an efficient algorithm for discovering statistically significant patterns in DNA sequences.

We present here a fast and sensitive method designed to isolate short nucleotide sequences which have non-random statistical properties and may thus be biologically active. It is based on a first order Markov analysis and allows us to detect statistically significant sequence motifs from six to ten nucleotides long which are significantly shared (or avoided) in the sequences under investigation. This method has been tested on a set of 521 sequences extracted from the Eukaryotic Promoter Database (2). Our results demonstrate the accuracy and the efficiency of the method in that the sequence motifs which are known to act as eukaryotic promoters, such as the TATA-box and the CAAT-box, were clearly identified. In addition we have found other statistically significant motifs, the biological roles of which are yet to be clarified.

Algorithms

Mitochondrial DNA detection and copy number determination in the spermatozoa of the sea urchin Arbacia lixula.

The Polymerase Chain reaction technique has been used in order to detect and amplify a specific region of mtDNA, in a total DNA preparation extracted from the sperm of the sea urchin Arbacia lixula. The amplified fragment is the D-loop region which hybridizes with the homologous region extracted from the egg mtDNA. The results demonstrate that mtDNA is present in sperm cell, and, since the replication origin is present it is potentially able to replicate in the zygote. Furthermore, the technique used allowed us to estimate mtDNA copy number in sea urchin sperm, which has never been done before. Our results are that sea urchin sperm cell contains between 4 and 28 mtDNA molecules.

Animals

Transcription mapping of the Ori L region reveals novel precursors of mature RNA species and antisense RNAs in rat mitochondrial genome.

We have identified new transcripts in the region surrounding the L-strand replication origin (Ori L) of rat liver mitochondrial DNA. In particular, we have detected previously unidentified intermediates of RNA processing on both the heavy and the light strands, such as precursors of the ND2 mRNA plus the Trp-tRNA and precursors of the tRNAs clustered in the Ori L region. This indicates that the mechanism of RNA processing in mitochondria proceeds step-wise producing a variety of precursors of the mature forms. The other striking finding is the detection of antisense RNA species in the region of L-strand replication. Since a variety of antisense transcripts were also found in the D-loop region of rat mitochondrial DNA, we suggest that they might play a regulatory role in the replication and expression of the mitochondrial genome.

Animals

A simple method for global sequence comparison.

A simple method of sequence comparison, based on a correlation analysis of oligonucleotide frequency distributions, is here shown to be a reliable test of overall sequence similarity. The method does not involve sequence alignment procedures and permits the rapid screening of large amounts of sequence data. It identifies those sequences which deserve more careful analysis of sequence similarity at the level of resolution of the single nucleotide. It uses observed quantities only and does not involve the adoption of any theoretical model.

Algorithms

A statistical method for detecting regions with different evolutionary dynamics in multialigned sequences.

We describe a stochastic method for tracing the evolutionary pattern of multialigned sequences. This method allows us to detect gene regions with distinct evolutionary dynamics, e.g., regions that significantly deviate from the expected behavior. Accurate detection of hypervariable or hyperconstrained regions may provide useful information on the structure/function relationship of biosequences. This information can help localize functional constraints. In addition, the selection of distinct evolutionary dynamics may assist in the correct use of biosequences as reliable molecular clocks.

Animals

The evolution of the mitochondrial D-loop region and the origin of modern man.

The origin of modern man is a highly debated issue that has recently been tackled by using mitochondrial DNA sequences. The limited genetic variability of human mtDNA has been explained in terms of a recent common genetic ancestry, thus implying that all modern-population mtDNAs originated from a single woman who lived in Africa less than 0.2 Mya. This divergence time is based on both the estimation of the rate of mtDNA change and its calibration date. Because different estimates of the rate of mtDNA evolution can completely change the scenario of the origin of modern man, we have reanalyzed the available mitochondrial sequence data by using an improved version of the statistical model, the "Markov clock," devised in our laboratory. Our analysis supports the African origin of modern man, but we found that the ancestral female from which all extant human mtDNAs originated lived in a time span of 0.3-0.8 Mya. Pushing back the date of the deepest root of the human implies that the earliest divergence would have been in the Homo erectus population.

Animals

Mitochondrial DNA in the sea urchin Arbacia lixula: nucleotide sequence differences between two polymorphic molecules indicate asymmetry of mutations.

Two polymorphic forms of mitochondrial DNA (mtDNA) extracted from Arbacia lixula eggs were cloned and the nucleotide sequences of specific regions determined. A comparison of the sequences of the sense strand of the two molecules demonstrates that all the differences are transitions and only of the A----G type. A change such as G----A (or A----G) on the sense mtDNA strand results from either a direct G----A (or A----G) mutation on that strand or a C----T (or T----C) on the complementary strand. None of the C----T (or T----C) changes were detected on the sense strand, which implies that the A----G mutation bias on the sense strand is not reversed for the other strand. Our observation indicates the existence of mechanisms acting asymmetrically on the two mtDNA strands, possibly during mtDNA replication.

Amino Acid Sequence

Glutamine synthetase gene evolution: a good molecular clock.

Glutamine synthetase (EC 6.3.1.2) gene evolution in various animals, plants, and bacteria was evaluated by a general stationary Markov model. The evolutionary process proved to be unexpectedly regular even for a time span as long as that between the divergence of prokaryotes from eukaryotes. This enabled us to draw phylogenetic trees for species whose phylogeny cannot be easily reconstructed from the fossil record. Our calculation of the times of divergence of the various organelle-specific enzymes led us to hypothesize that the pea and bean chloroplast genes for these enzymes originated from the duplication of nuclear genes as a result of the different metabolic needs of the various species. Our data indicate that the duplication of plastid glutamine synthetase genes occurred long after the endosymbiotic events that produced the organelles themselves.

Amino Acid Sequence

Evolutionary analysis of the nucleus-encoded subunits of mammalian cytochrome c oxidase.

The cytochrome c oxidase enzyme complex of eukaryotes is made up of three mitochondrial-coded subunits and a variable number of nuclear-coded subunits. Some nuclear-coded subunits are present in multiple forms and probably perform a tissue- or development-specific function. A detailed evolutionary analysis of the cytochrome c oxidase subunits that have been sequenced to date is reported here. We have found that gene duplication events from which the liver and heart isoforms of rat subunits VIa and subunit VIII originated can both be dated at about 240 +/- 90 million years ago, long before the radiation of mammalian lineages. Sequence divergence between the processed-type pseudogenes for the subunits IV, VIc and VIII have been estimated. Our results indicate that they arose fairly recently, thus suggesting that retroposition is a continuing process. We show that the rate of silent substitution in mitochondrial-coded subunits is 5-10 times higher than in nuclear-coded subunits; on the other hand replacement rates, although differing from gene to gene, are roughly of the same order of magnitude in both nuclear and mitochondrial genes. In the case of most of the nuclear-coded proteins we observed a slightly greater similarity between rats and cow, which agrees with the data obtained for mitochondrial-coded subunits.

Amino Acid Sequence

The main regulatory region of mammalian mitochondrial DNA: structure-function model and evolutionary pattern.

The evolution of the main regulatory region (D-loop) of the mammalian mitochondrial genome was analyzed by comparing the sequences of eight mammalian species: human, common chimpanzee, pygmy chimpanzee, dolphin, cow, rat, mouse, and rabbit. The best alignment of the sequences was obtained by optimization of the sequence similarities common to all these species. The two peripheral left and right D-loop domains, which contain the main regulatory elements so far discovered, evolved rapidly in a species-specific manner generating heterogeneity in both length and base composition. They are prone to the insertion and deletion of elements and to the generation of short repeats by replication slippage. However, the preservation of some sequence blocks and similar cloverleaf-like structures in these regions, indicates a basic similarity in the regulatory mechanisms of the mitochondrial genome in all mammalian species. We found, particularly in the right domain, significant similarities to the telomeric sequences of the mitochondrial (mt) and nuclear DNA of Tetrahymena thermophila. These sequences may be interpreted as relics of telomeres present in ancestral linear forms of mtDNA or may simply represent efficient templates of RNA primase-like enzymes. Due to their peculiar evolution, the two peripheral domains cannot be used to estimate in a quantitative way the genetic distances between mammalian species. On the other hand the central domain, highly conserved during evolution, behaves as a good molecular clock. Reliable estimates of the times of divergence between closely and distantly related species were obtained from the central domain using a Markov model and assuming nonhomogeneous evolution of nucleotide sites.

Animals

The branching order of mammals: phylogenetic trees inferred from nuclear and mitochondrial molecular data.

In order to clarify some controversial phylogenies such as those regarding the triplet of human, rodent, and cow and the evolutionary position of Lagomorpha with respect to other mammals, we have analyzed both nuclear and mitochondrial genes using the stationary Markov model developed in our laboratory. We found that the two sets of genes give different results. In particular the mitochondrial tree showed rabbit linked first to rodents and the rabbit-rodents branch linked to artiodactyls with human as the outgroup. The most favorite nuclear tree showed human linked first to artiodactyls and the human-artiodactyls branch linked to rabbit with rodents as the outgroup. The obvious questions, (1) which tree is the correct one, or (2) both trees can be incorrect, and (3) how can we explain such an evolutionary pattern, are discussed on the basis of our limited knowledge of factors that influence the clocklike behavior of biological macromolecules.

Animals

Mitochondrial DNA in the sea urchin Arbacia lixula: evolutionary inferences from nucleotide sequence analysis.

From the stirodont Arbacia lixula we determined the sequence of 5,127 nucleotides of mitochondrial DNA (mtDNA) encompassing 18 tRNAs, two complete coding genes, parts of three other coding genes, and part of the 12S ribosomal RNA (rRNA). The sequence confirms that the organization of mtDNA is conserved within echinoids. Furthermore, it underlines the following peculiar features of sea urchin mtDNA: the clustering of tRNAs, the short noncoding regulatory sequence, and the separation by the ND1 and ND2 genes of the two rRNA genes. Comparison with the orthologous sequences from the camarodont species Paracentrotus lividus and Strongylocentrotus purpuratus revealed that (1) echinoids have an extra piece on the amino terminus of the ND5 gene that is probably the remnant of an old leucine tRNA gene; (2) third-position codon nucleotide usage has diverged between A. lixula and the camarodont species to a significant extent, implying different directional mutational pressures; and (3) the stirodont-camarodont divergence occurred twice as long ago as did the P. lividus-S. purpuratus divergence.

Amino Acid Sequence

Direct evidence that restriction endonucleases may under estimate the degree of divergence between molecules.

We studied two polymorphic forms of mtDNA extracted from A. lixula eggs. In order to compare and to quantitate the variability, we sequenced specific regions of the two molecules. In this way, we obtained a precise measurement of the variability within two haplotypes. We also obtained a direct demonstration that some differences in nucleotide sequence can escape detection when restriction endonuclease analysis is used. Our results underline the unreliability of the use of restriction mapping to estimate divergence between relatively short and closely related DNA sequences.

Animals

The complete and symmetric transcription of the main non coding region of rat mitochondrial genome: in vivo mapping of heavy and light transcripts.

The experiments here reported demonstrate that the main non-coding region of rat mitochondrial DNA is symmetrically transcribed. We have identified stable heavy and light transcripts, whose pattern is rather complex, in the D-loop region of rat mitochondrial DNA. Their relative concentrations have been determined. We detected heavy transcripts which encompass the whole D-loop and more abundant heavy RNA species which we interpreted as transcripts terminating downstream of the 3' end of the last coded gene (Thr-tRNA). The processed heavy RNA species contain polyA, suggesting a strict association between cleavage and polyadenylation. The pattern of light transcripts shows a long RNA, which, starting from the light strand promoter, covers the whole segment, and shorter RNA species which seems to be actively processed at the level of the conserved sequence boxes, probably acting as primers. The symmetric transcription of the D-loop containing region of rat mitochondrial DNA, and in particular the presence of stable transcripts complementary to the putative RNA primers, suggest that mechanisms mediated by interaction between complementary transcripts (antisense RNAs) might play a role in the regulation of mitochondrial DNA replication and expression.

Amino Acid Sequence

Sequence-dependent DNA curvature: conformational signal present in the main regulatory region of the rat mitochondrial genome.

Theoretical analysis and experimental approaches by gel electrophoresis in retarding conditions allowed us to identify the presence of an intrinsic bending in the D-loop containing region of the rat mitochondrial genome. The curvature was located in the right domain of the sequence analyzed, between the origin of replication of the heavy strand and its promoter. The preliminary evidence of a specific recognition of the bent DNA with mitochondrial matrix proteins suggests a probable role of this DNA conformation in the duplication and/or expression of the mammalian mitochondrial genome.

Animals

The complete nucleotide sequence, gene organization, and genetic code of the mitochondrial genome of Paracentrotus lividus.

The 15,697-nucleotide sequence of Paracentrotus lividus mitochondrial DNA is reported. This genome codes for 2 rRNAs, 22 tRNAs, and 12 mRNAs which specify 13 subunits of the mitochondrial inner membrane respiratory complexes. The gene arrangement differs from that of other animal species. The two ribosomal genes 16 S and 12 S are separated by a stretch of about 3.3 kilobase pairs which contains the ND1 and ND2 genes and a cluster of 15 tRNA genes. The ND4L coding sequence is not contained in the ND4 mRNA but has its own mRNA which maps between the tRNA(Arg) and the Co II genes. The main noncoding region, located in the tRNA gene cluster, is only 132 nucleotides long, but contains sequences homologous to the mammalian displacement loop. Other short noncoding sequences are interspersed in the genome: they contain a conserved AT consensus which probably has a role in transcription or RNA processing. As regards the mitochondrial genetic code, the codons AGA and AGG specify serine and are recognized by a tRNA with a GCU anticodon, whereas AUA and AAA code for isoleucine and asparagine rather than for methionine and lysine. Except for ND4L which starts with AUC and ATPase 8 which starts with GUG, AUG is used as the initiation codon. In 11 out of 13 cases the genes terminate with the canonical stop codons UAA or UAG. These observations suggest that during invertebrate evolution each lineage developed its own mechanism of mitochondrial DNA replication and transcription and of RNA processing and translation.

Amino Acid Sequence