Search PubMed⌕ Search

Biomedical subjects

Leonid Kruglyak

Publications and source records attributed to Leonid Kruglyak.

16 recordsLinked to original sources

Extensive and breed-specific linkage disequilibrium in Canis familiaris.

The 156 breeds of registered dogs in the United States offer a unique opportunity to map genes important in disease susceptibility, morphology, and behavior. Linkage disequilibrium (LD) is of current interest for its application in whole genome association mapping, since the extent of LD determines the feasibility of such studies. We have measured LD at five genomic intervals, each 5 Mb in length and composed of five clusters of sequence variants spaced 800 kb-1.6 Mb apart. These intervals are located on canine chromosomes 1, 2, 3, 34, and 37, and none is under obvious selective pressure. Approximately 20 unrelated dogs were assayed from each of five breeds: Akita, Bernese Mountain Dog, Golden Retriever, Labrador Retriever, and Pekingese. At each genomic interval, SNPs and indels were discovered and typed by resequencing. Strikingly, LD in canines is much more extensive than in humans: D' falls to 0.5 at 400-700 kb in Golden Retriever and Labrador Retriever, 2.4 Mb in Akita, and 3-3.2 Mb in Bernese Mountain Dog and Pekingese. LD in dog breeds is up to 100x more extensive than in humans, suggesting that a correspondingly smaller number of markers will be required for association mapping studies in dogs compared to humans. We also report low haplotype diversity within regions of high LD, with 80% of chromosomes in a breed carrying two to four haplotypes, as well as a high degree of haplotype sharing among breeds.

Animals↗

Population history and natural selection shape patterns of genetic variation in 132 genes.

Identifying regions of the human genome that have been targets of natural selection will provide important insights into human evolutionary history and may facilitate the identification of complex disease genes. Although the signature that natural selection imparts on DNA sequence variation is difficult to disentangle from the effects of neutral processes such as population demographic history, selective and demographic forces can be distinguished by analyzing multiple loci dispersed throughout the genome. We studied the molecular evolution of 132 genes by comprehensively resequencing them in 24 African-Americans and 23 European-Americans. We developed a rigorous computational approach for taking into account multiple hypothesis tests and demographic history and found that while many apparent selective events can instead be explained by demography, there is also strong evidence for positive or balancing selection at eight genes in the European-American population, but none in the African-American population. Our results suggest that the migration of modern humans out of Africa into new environments was accompanied by genetic adaptations to emergent selective forces. In addition, a region containing four contiguous genes on Chromosome 7 showed striking evidence of a recent selective sweep in European-Americans. More generally, our results have important implications for mapping genes underlying complex human diseases.

Black People↗

Sequence-based linkage analysis.

The rapid decrease in the cost of DNA sequencing will enable its use for novel applications. Here, we investigate the use of DNA sequencing for simultaneous discovery and genotyping of polymorphisms in family linkage studies. In the proposed approach, short contiguous segments of genomic DNA, regularly spaced across the genome, are resequenced in each pedigree member, and all sequence polymorphisms discovered within a pedigree are used as genetic markers. We use computer simulations consistent with observed human sequence diversity to show that segments of 500-1,000 base pairs, spaced at intervals of 1-2 Mb across the genome, provide linkage information that equals or exceeds that of traditional marker-based approaches. We validate these results experimentally by implementing the sequence-based linkage approach for chromosome 19 in CEPH pedigrees.

Chromosome Mapping↗

Mapping complex disease loci in whole-genome association studies.

Identification of the genetic polymorphisms that contribute to susceptibility for common diseases such as type 2 diabetes and schizophrenia will aid in the development of diagnostics and therapeutics. Previous studies have focused on the technique of genetic linkage, but new technologies and experimental resources make whole-genome association studies more feasible. Association studies of this type have good prospects for dissecting the genetics of common disease, but they currently face a number of challenges, including problems with multiple testing and study design, definition of intermediate phenotypes and interaction between polymorphisms.

Disease↗

Genetic structure of the purebred domestic dog.

We used molecular markers to study genetic relationships in a diverse collection of 85 domestic dog breeds. Differences among breeds accounted for approximately 30% of genetic variation. Microsatellite genotypes were used to correctly assign 99% of individual dogs to breeds. Phylogenetic analysis separated several breeds with ancient origins from the remaining breeds with modern European origins. We identified four genetic clusters, which predominantly contained breeds with similar geographic origin, morphology, or role in human activities. These results provide a genetic classification of dog breeds and will aid studies of the genetics of phenotypic breed differences.

Algorithms↗

Haplotype diversity across 100 candidate genes for inflammation, lipid metabolism, and blood pressure regulation in two populations.

Recent studies have suggested that a significant fraction of the human genome is contained in blocks of strong linkage disequilibrium, ranging from ~5 to >100 kb in length, and that within these blocks a few common haplotypes may account for >90% of the observed haplotypes. Furthermore, previous studies have suggested that common haplotypes in candidate genes are generally shared across populations and represent the majority of chromosomes in each population. The conclusions drawn from these preliminary studies, however, are based on an incomplete knowledge of the variation in the regions examined. To bridge this gap in knowledge, we have completely resequenced 100 candidate genes in a population of African descent and one of European descent. Although these genes have been well studied because of their medical importance, we demonstrate that a large amount of sequence variation has not yet been described. We also report that the average number of inferred haplotypes per gene, when complete data is used, is higher than in previous reports and that the number and proportion of all haplotypes represented by common haplotypes per gene is variable. Furthermore, we demonstrate that haplotypes shared between the two populations constitute only a fraction of the total number of haplotypes observed and that these shared haplotypes represent fewer of the African-descent chromosomes than was expected from previous studies. Finally, we show that restricting variation discovery to coding regions does not adequately describe all common haplotypes or the true haplotype block structure observed when all common variation is used to infer haplotypes. These data, derived from complete knowledge of genetic variation in these genes, suggest that the haplotype architecture of candidate genes across the human genome is more complex than previously suggested, with important implications for candidate gene and genomewide association studies.

Africa↗

Selecting a maximally informative set of single-nucleotide polymorphisms for association analyses using linkage disequilibrium.

Common genetic polymorphisms may explain a portion of the heritable risk for common diseases. Within candidate genes, the number of common polymorphisms is finite, but direct assay of all existing common polymorphism is inefficient, because genotypes at many of these sites are strongly correlated. Thus, it is not necessary to assay all common variants if the patterns of allelic association between common variants can be described. We have developed an algorithm to select the maximally informative set of common single-nucleotide polymorphisms (tagSNPs) to assay in candidate-gene association studies, such that all known common polymorphisms either are directly assayed or exceed a threshold level of association with a tagSNP. The algorithm is based on the r(2) linkage disequilibrium (LD) statistic, because r(2) is directly related to statistical power to detect disease associations with unassayed sites. We show that, at a relatively stringent r(2) threshold (r2>0.8), the LD-selected tagSNPs resolve >80% of all haplotypes across a set of 100 candidate genes, regardless of recombination, and tag specific haplotypes and clades of related haplotypes in nonrecombinant regions. Thus, if the patterns of common variation are described for a candidate gene, analysis of the tagSNP set can comprehensively interrogate for main effects from common functional variation. We demonstrate that, although common variation tends to be shared between populations, tagSNPs should be selected separately for populations with different ancestries.

Algorithms↗

Trans-acting regulatory variation in Saccharomyces cerevisiae and the role of transcription factors.

Natural genetic variation can cause significant differences in gene expression, but little is known about the polymorphisms that affect gene regulation. We analyzed regulatory variation in a cross between laboratory and wild strains of Saccharomyces cerevisiae. Clustering and linkage analysis defined groups of coregulated genes and the loci involved in their regulation. Most expression differences mapped to trans-acting loci. Positional cloning and functional assays showed that polymorphisms in GPA1 and AMN1 affect expression of genes involved in pheromone response and daughter cell separation, respectively. We also asked whether particular classes of genes were more likely to contain trans-regulatory polymorphisms. Notably, transcription factors showed no enrichment, and trans-regulatory variation seems to be broadly dispersed across classes of genes with different molecular functions.

Amino Acid Sequence↗

A 3.9-centimorgan-resolution human single-nucleotide polymorphism linkage map and screening set.

Recent advances in technologies for high-throughout single-nucleotide polymorphism (SNP)-based genotyping have improved efficiency and cost so that it is now becoming reasonable to consider the use of SNPs for genomewide linkage analysis. However, a suitable screening set of SNPs and a corresponding linkage map have yet to be described. The SNP maps described here fill this void and provide a resource for fast genome scanning for disease genes. We have evaluated 6,297 SNPs in a diversity panel composed of European Americans, African Americans, and Asians. The markers were assessed for assay robustness, suitable allele frequencies, and informativeness of multi-SNP clusters. Individuals from 56 Centre d'Etude du Polymorphisme Humain pedigrees, with >770 potentially informative meioses altogether, were genotyped with a subset of 2,988 SNPs, for map construction. Extensive genotyping-error analysis was performed, and the resulting SNP linkage map has an average map resolution of 3.9 cM, with map positions containing either a single SNP or several tightly linked SNPs. The order of markers on this map compares favorably with several other linkage and physical maps. We compared map distances between the SNP linkage map and the interpolated SNP linkage map constructed by the deCode Genetics group. We also evaluated cM/Mb distance ratios in females and males, along each chromosome, showing broadly defined regions of increased and decreased rates of recombination. Evaluations indicate that this SNP screening set is more informative than the Marshfield Clinic's commonly used microsatellite-based screening set.

Alleles↗

Additional SNPs and linkage-disequilibrium analyses are necessary for whole-genome association studies in humans.

More than 5 million single-nucleotide polymorphisms (SNPs) with minor-allele frequency greater than 10% are expected to exist in the human genome. Some of these SNPs may be associated with risk of developing common diseases. To assess the power of currently available SNPs to detect such associations, we resequenced 50 genes in two ethnic samples and measured patterns of linkage disequilibrium between the subset of SNPs reported in dbSNP and the complete set of common SNPs. Our results suggest that using all 2.7 million SNPs currently in the database would detect nearly 80% of all common SNPs in European populations but only 50% of those common in the African American population and that efficient selection of a minimal subset of SNPs for use in association studies requires measurement of allele frequency and linkage disequilibrium relationships for all SNPs in dbSNP.

Alleles↗

Genetic loci affecting resistance to human malaria parasites in a West African mosquito vector population.

Successful propagation of the malaria parasite Plasmodium falciparum within a susceptible mosquito vector is a prerequisite for the transmission of malaria. A field-based genetic analysis of the major human malaria vector, Anopheles gambiae, has revealed natural factors that reduce the transmission of P. falciparum. Differences in P. falciparum oocyst numbers between mosquito isofemale families fed on the same infected blood indicated a large genetic component affecting resistance to the parasite, and genome-wide scanning in pedigrees of wild mosquitoes detected segregating resistance alleles. The apparently high natural frequency of resistance alleles suggests that malaria parasites (or a similar pathogen) exert a significant selective pressure on vector populations.

Alleles↗

Genetic dissection of transcriptional regulation in budding yeast.

To begin to understand the genetic architecture of natural variation in gene expression, we carried out genetic linkage analysis of genomewide expression patterns in a cross between a laboratory strain and a wild strain of Saccharomyces cerevisiae. Over 1500 genes were differentially expressed between the parent strains. Expression levels of 570 genes were linked to one or more different loci, with most expression levels showing complex inheritance patterns. The loci detected by linkage fell largely into two categories: cis-acting modulators of single genes and trans-acting modulators of many genes. We found eight such trans-acting loci, each affecting the expression of a group of 7 to 94 genes of related function.

Chromosome Mapping↗

A new susceptibility locus for autosomal dominant pancreatic cancer maps to chromosome 4q32-34.

Pancreatic cancer is the fifth leading cause of cancer death in the United States. Nearly every person diagnosed with pancreatic cancer will die from it, usually in <6 mo. Familial clustering of pancreatic cancers is commonly recognized, with an autosomal dominant inheritance pattern in approximately 10% of all cases. However, the late age at disease onset and rapid demise of affected individuals markedly hamper collection of biological samples. We report a genetic linkage scan of family X with an autosomal dominant pancreatic cancer with early onset and high penetrance. For the study of this family, we have developed an endoscopic surveillance program that allows the early detection of cancer and its precursor, before family members have died of the disease. In a genomewide screening of 373 microsatellite markers, we found significant linkage (maximum LOD score 4.56 in two-point analysis and 5.36 in three-point analysis) on chromosome 4q32-34, providing evidence for a major locus for pancreatic cancer.

Adult↗

Patterns of linkage disequilibrium in the human genome.

Particular alleles at neighbouring loci tend to be co-inherited. For tightly linked loci, this might lead to associations between alleles in the population a property known as linkage disequilibrium (LD). LD has recently become the focus of intense study in the hope that it might facilitate the mapping of complex disease loci through whole-genome association studies. This approach depends crucially on the patterns of LD in the human genome. In this review, we draw on empirical studies in humans and Drosophila, as well as simulation studies, to assess the current state of knowledge about patterns of LD, and consider the implications for the use of LD as a mapping tool.

Animals↗