Search PubMed⌕ Search

Biomedical subjects

J Gordon Burleigh

Publications and source records attributed to J Gordon Burleigh.

6 recordsLinked to original sources

Supertree bootstrapping methods for assessing phylogenetic variation among genes in genome-scale data sets.

Nonparamtric bootstrapping methods may be useful for assessing confidence in a supertree inference. We examined the performance of two supertree bootstrapping methods on four published data sets that each include sequence data from more than 100 genes. In "input tree bootstrapping," input gene trees are sampled with replacement and then combined in replicate supertree analyses; in "stratified bootstrapping," trees from each gene's separate (conventional) bootstrap tree set are sampled randomly with replacement and then combined. Generally, support values from both supertree bootstrap methods were similar or slightly lower than corresponding bootstrap values from a total evidence, or supermatrix, analysis. Yet, supertree bootstrap support also exceeded supermatrix bootstrap support for a number of clades. There was little overall difference in support scores between the input tree and stratified bootstrapping methods. Results from supertree bootstrapping methods, when compared to results from corresponding supermatrix bootstrapping, may provide insights into patterns of variation among genes in genome-scale data sets.

Algorithms↗

Identifying optimal incomplete phylogenetic data sets from sequence databases.

We introduce a new method for identifying optimal incomplete data sets from large sequence databases based on the graph theoretic concept of alpha-quasi-bicliques. The quasi-biclique method searches large sequence databases to identify useful phylogenetic data sets with a specified amount of missing data while maintaining the necessary amount of overlap among genes and taxa. The utility of the quasi-biclique method is demonstrated on large simulated sequence databases and on a data set of green plant sequences from GenBank. The quasi-biclique method greatly increases the taxon and gene sampling in the data sets while adding only a limited amount of missing data. Furthermore, under the conditions of the simulation, data sets with a limited amount of missing data often produce topologies nearly as accurate as those built from complete data sets. The quasi-biclique method will be an effective tool for exploiting sequence databases for phylogenetic information and also may help identify critical sequences needed to build large phylogenetic data sets.

Base Sequence↗

Covarion structure in plastid genome evolution: a new statistical test.

Covarion models of molecular evolution allow the rate of evolution of a site to vary through time. There are few simple and effective tests for covarion evolution, and consequently, little is known about the presence of covarion processes in molecular evolution. We describe two new tests for covarion evolution and demonstrate with simulations that they perform well under a wide range of conditions. A survey of covarion evolution in sequenced plastid genomes found evidence of covarion drift in at least 26 out of 57 genes. Covarion evolution is most evident in first and second codon positions of the plastid genes, and there is no evidence of covarion evolution in third codon positions. Therefore, the significant covarion tests are likely due to changes in the selective constraints of amino acids. The frequency of covarion evolution within the plastid genome suggests that covarion processes of evolution were important in generating the observed patterns of sequence variation among plastid genomes.

Codon↗

Prospects for building the tree of life from large sequence databases.

We assess the phylogenetic potential of approximately 300,000 protein sequences sampled from Swiss-Prot and GenBank. Although only a small subset of these data was potentially phylogenetically informative, this subset retained a substantial fraction of the original taxonomic diversity. Sampling biases in the databases necessitate building phylogenetic data sets that have large numbers of missing entries. However, an analysis of two "supermatrices" suggests that even data sets with as much as 92% missing data can provide insights into broad sections of the tree of life.

Animals↗

Performance of flip supertree construction with a heuristic algorithm.

Supertree methods are used to assemble separate phylogenetic trees with shared taxa into larger trees (supertrees) in an effort to construct more comprehensive phylogenetic hypotheses. In spite of much recent interest in supertrees, there are still few methods for supertree construction. The flip supertree problem is an error correction approach that seeks to find a minimum number of changes (flips) to the matrix representation of the set of input trees to resolve their incompatibilities. A previous flip supertree algorithm was limited to finding exact solutions and was only feasible for small input trees. We developed a heuristic algorithm for the flip supertree problem suitable for much larger input trees. We used a series of 48- and 96-taxon simulations to compare supertrees constructed with the flip supertree heuristic algorithm with supertrees constructed using other approaches, including MinCut (MC), modified MC (MMC), and matrix representation with parsimony (MRP). Flip supertrees are generally far more accurate than supertrees constructed using MC or MMC algorithms and are at least as accurate as supertrees built with MRP. The flip supertree method is therefore a viable alternative to other supertree methods when the number of taxa is large.

Algorithms↗

Adaptive evolution in the photosensory domain of phytochrome A in early angiosperms.

Flowering plant diversity now far exceeds the combined diversity of all other plant groups. Recently identified extant remnants of the earliest-diverging lines suggest that the first angiosperms may have lived in shady, disturbed, and moist understory habitats, and that the aquatic habit also arose early. This would have required the capacity to begin life in dimly lit environments. If so, evolution in light-sensing mechanisms may have been crucial to their success. The photoreceptor phytochrome A is unique among angiosperm phytochromes in its capacity to serve a transient role under conditions where an extremely high sensitivity is required. We present evidence of altered functional constraints between phytochrome A (PHYA) and its paralog, PHYC. Tests for selection suggest that an elevation in nonsynonymous rates resulted from an episode of selection along the branch leading to all angiosperm PHYA sequences. Most nucleotide sites (95%) are selectively constrained, and the ratio of nonsynonymous to synonymous substitutions on branches within the PHYA clade does not differ from the ratio on the branches in the PHYC clade. Thus, positive selection at a handful of sites, rather than relaxation of selective constraints, apparently has played a major role in the evolution of the photosensory domain of phytochrome A. The episode of selection occurred very early in the history of flowering plants, suggesting that innovation in phyA may have given the first angiosperms some adaptive advantage.

Evolution, Molecular↗